Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Practical, Colab-friendly notebooks for fine-tuning and running audio AI models
| Date | Stars |
|---|---|
| 2026-07-31 | 420 |
| 2026-08-04 | 420 |
| 2026-08-06 | 420 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
<p align="center"> <img src="image.png" alt="Smol Audio" /> </p> # Smol Audio 🔊 Practical notebooks for shrinking, optimizing, and customizing audio AI models with the Hugging Face ecosystem. ### Latest examples - Inference with Perception Encoder for Audio-Video (PE-AV) - Fine-tune Audio Flamingo 3 - Granite Speech 4.0 1b ASR > [!NOTE] > GitHub doesn't always render notebooks well. If you have trouble viewing them, try opening in Colab using the links below. | Category | Notebook | Description | |----------|----------|-------------| | ASR Fine-tuning | [Fine-tune Whisper](Fine_tune_Whisper.ipynb) | Fine-tune Whisper on a custom language/domain using transformers + datasets | | ASR Fine-tuning | [Fine-tune Granite Speech Italian](Fine_tune_Granite_Speech_Italian.ipynb) | Fine-tune IBM Granite Speech for Italian ASR with the YODAS-Granary dataset | | Audio Captioning | [Fine-tune Audio Flamingo 3](Fine_tune_Audio_Flamingo_3.ipynb) | Fine-tune Audio Flamingo 3 for audio captioning (full + LoRA) | | ASR Fine-tuning | [Fine-tune Parakeet](Fine_tune_Parakeet.ipynb) | Fine-tune NVIDIA Parakeet CTC for speech recognition (full + LoRA) | | ASR Fine-tuning | [Fine-tune Voxtral ASR](Fine_tune_Voxtral_ASR.ipynb) | Fine-tune Voxtral for ASR with prompt masking (full + LoRA) | | Multimodal | [Inference with PE-AV-Base](Inference_PE_AV_Base.ipynb) | Zero-shot video classification and audio↔text retrieval (AudioCaps) with Meta's Perception Encoder for Audio-Video |
Excerpt of 1,483 characters
Read on GitHubWould you bet a product on this? Bounded 0–100 and slow moving.
matched fp:e2d1cad02b642923, desc:fine-tuning, desc:fine tuning