Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Worlds first open-source real-time end-to-end spoken dialogue model with personalized voice cloning.
| Date | Stars |
|---|---|
| 2026-07-24 | 550 |
| 2026-07-25 | 550 |
| 2026-07-28 | 550 |
| 2026-07-30 | 550 |
| 2026-08-07 | 550 |
| 2026-08-15 | 551 |
| 2026-08-20 | 550 |
| 2026-08-31 | 550 |
| 2026-09-03 | 549 |
| 2026-09-04 | 549 |
| 2026-09-07 | 549 |
| 2026-09-09 | 550 |
| 2026-09-20 | 550 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
<div align="center"> # FlashLabs Chroma 1.0: A Real-Time End-to-End Spoken Dialogue Model with Personalized Voice Cloning </div> <br> <div align="center"> <img src="figures/logo.svg" alt="FlashLabs Chroma Logo" width="400px"/> <h3>🚀 Get Started with <a href="https://www.flashlabs.ai/flashai-voice-agents">FlashAI Voice Agents</a>!</h3> <p><strong>Production-ready voice AI solutions</strong> powered by Chroma | <strong>Open-source model</strong> for developers & researchers</p> [](https://www.flashlabs.ai/flashai-voice-agents) [](https://huggingface.co/FlashLabs/Chroma-4B) [](https://arxiv.org/abs/2601.11141) [](https://x.com/flashlabsdotai) [](https://www.linkedin.com/company/flashlabs-ai/) </div> ## 🎬 Demo <div align="center"> https://github.com/user-attachments/assets/3723c24d-d262-4c3e-88ee-34a16546359e <p><i>Watch our model in action</i></p> </div> ## Model Description **Chroma 1.0** is an advanced multimodal model developed by **[FlashLabs](https://flashlabs.ai)**. It is designed to understand and generate content across multiple modalities, including text and audio. As a virtual human model, Chroma possesses the ability to process auditory inputs and respond with both text and synthesized speech, enabling natural voice interactions. - **Model Type:** Multimodal Causal Language Model - **Developed by:** FlashLabs - **Language(s):** English - **License:** Apache-2.0 - **Model Architecture:** - **Reasoner:** Based on Qwen2.5-Omni-3B - **Backbone:** Based on Llama3 (16 layers, 2048 hidden size) - **Decoder:** Based on Llama3 (4 layers, 1024 hidden size) - **Codec:** Mimi (24kHz sampling rate) ## Model Architecture <img src="figures/architecture.png" alt="Model Architecture" width="800" /> ## Capabilities Chroma 1.0 is capable of: - **Speech Understanding:** Processing user audio input directly. - **Multimodal Generation:** Generating coherent text and speech responses simultaneously. - **Voice Cloning:** utilizing reference audio prompts to guide speech generation style. ## Usage ### Installation #### Requirements - Python 3.11 or higher - CUDA 12.6 or compatible version (for GPU support) #### Quick Install ```bash # Clone the repository git clone https://github.com/FlashLabs-AI-Corp/FlashLabs-Chroma.
Excerpt of 8,674 characters
Read on GitHub18
2
2
1
1
1
1
Would you bet a product on this? Bounded 0–100 and slow moving.
matched fp:4e29a7c13d675c36, topic:voice-cloning, desc:voice cloning, readme:voice cloning