Switch language한국어
Back to the list

smol-audio: A Colab-Friendly Notebook Collection for Fine-Tuning Whisper, Parakeet, Voxtral, Granite Speech, and Audio Flamingo 3

TL;DR AI

Key summary

2 min read
  1. Deep-unlearning released smol-audio, an Apache-2.0 collection of self-contained Jupyter notebooks for audio AI work.

  2. The notebooks cover fine-tuning ASR models, LoRA, CTC, prompt masking, and practical workflows for speech and audio tasks.

  3. They reproduce and adapt models including Whisper, Parakeet, Voxtral, Granite Speech, and Audio Flamingo 3.

  4. Designed to run in Colab, smol-audio makes modern audio model adaptation more transparent and accessible.

Read the original