Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Generate audiobooks from EPUBs, PDFs and text with synchronized captions.
| Date | Stars |
|---|---|
| 2026-07-24 | 5340 |
| 2026-07-25 | 5350 |
| 2026-07-28 | 5350 |
| 2026-07-30 | 5350 |
| 2026-08-06 | 5350 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
# abogen <img width="40px" title="abogen icon" src="https://raw.githubusercontent.com/denizsafak/abogen/refs/heads/main/abogen/assets/icon.ico" align="right" style="padding-left: 10px; padding-top:5px;"> [](https://github.com/denizsafak/abogen/actions) [](https://github.com/denizsafak/abogen/releases/latest) [](https://pypi.org/project/abogen/) [](https://github.com/denizsafak/abogen/releases/latest) [&color=blue)](https://pypi.org/project/abogen/) [](https://github.com/psf/black) [](https://opensource.org/licenses/MIT) <a href="https://trendshift.io/repositories/14433" target="_blank"><img src="https://trendshift.io/api/badge/repositories/14433" alt="denizsafak%2Fabogen | Trendshift" style="width: 250px; height: 55px;" width="250" height="55"/></a> Abogen is a powerful text-to-speech conversion tool that makes it easy to turn ePub, PDF, text, markdown, or subtitle files into high-quality audio with matching subtitles in seconds. Use it for audiobooks, voiceovers for Instagram, YouTube, TikTok, or any project that needs natural-sounding text-to-speech, using [Kokoro-82M](https://huggingface.co/hexgrad/Kokoro-82M). <img title="Abogen Main" src='https://raw.githubusercontent.com/denizsafak/abogen/refs/heads/main/demo/abogen.png' width="380"> <img title="Abogen Processing" src='https://raw.githubusercontent.com/denizsafak/abogen/refs/heads/main/demo/abogen2.png' width="380"> ## Demo https://github.com/user-attachments/assets/094ba3df-7d66-494a-bc31-0e4b41d0b865 > This demo was generated in just 5 seconds, producing ∼1 minute of audio with perfectly synced subtitles. To create a similar video, see [the demo guide](https://github.com/denizsafak/abogen/tree/main/demo). ## `How to install?` <a href="https://pypi.org/project/abogen/" target="_blank"><img src="https://img.shields.io/pypi/pyversions/abogen" alt="Abogen Compatible PyPi Python Versions" align="right" style="margin-top:6px;"></a> ### `Windows` Go to [espeak-ng latest release](https://github.com/espeak-ng/espeak-ng/releases/latest) download and run the *.msi file. #### <b>OPTION 1: Install using script</b> 1. [Download](https://github.com/denizsafak/abogen/archive/refs/heads/main.zip) the repository 2. Extract the ZIP file 3. Run `WINDOWS_INSTALL.bat` by double-clicking it This method handles everything automatically - installing all dependencies including CUDA in a self-contained environment without requiring a separate Python installation. (You still need to install [espeak-ng](https://github.com/espeak-ng/espeak-ng/releases/latest).) > [!NOTE] > You don't need to install Python separately. The script will install Python automatically. #### <b>OPTION 2: Install using uv</b> First, [install uv](https://docs.astral.sh/uv/getting-started/installation/) if you haven't already. ```bash # For NVIDIA GPUs (CUDA 12.8) - Recommended uv tool install --python 3.12 abogen[cuda] --extra-index-url https://download.pytorch.org/whl/cu128 --index-strategy unsafe-best-match # For NVIDIA GPUs (CUDA 12.6) - Older drivers uv tool install --python 3.12 abogen[cuda126] --extra-index-url https://download.pytorch.org/whl/cu126 --index-strategy unsafe-best-match # For NVIDIA GPUs (CUDA 13.0) - Newer drivers uv tool install --python 3.12 abogen[cuda130] --extra-index-url https://download.pytorch.org/whl/cu130 --index-strategy unsafe-best-match # For AMD GPUs or without GPU - If you have AMD GPU, you need to use Lin
Excerpt of 43,497 characters
Read on GitHub345
161
33
9
3
2
1
1
1
1
1
Vladimir Sol
1
Dong Xia · Shanghai Jiao Tong University
1
1
1
1
Would you bet a product on this? Bounded 0–100 and slow moving.
matched fp:296ee4dd314f7308, topic:text-to-speech, topic:tts, topic:speech-synthesis