BandarLabs/gitpodcast
quality grade C, 54 out of 100Convert any git repository into an engaging podcast
- stars
- 810
- stars gained this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Speech recognition, text-to-speech, voice cloning, music generation and audio processing.
Signals: speech-recognition, text-to-speech, tts, stt, asr, whisper, voice-cloning, speech-synthesis
915 results
Convert any git repository into an engaging podcast
Voxtral ASR & TTS running natively and in the browser. A Rust implementation of Mistral's Voxtral mini realtime ASR / TTS using the Burn ML framework
The simplest and lowest-cost AI integration solution. If you like this project, please give it a Star~ | 最简单、最低成本的AI接入方案。喜欢本项目的话点个 Star 吧~
Transform PDFs into AI podcasts for engaging on-the-go audio content.
Suno AI's Bark model in C/C++ for fast text-to-speech generation
Make Azure natural TTS voices accessible to any SAPI 5-compatible application.
A user-friendly toolkit for voice recgonition/transcription/conversion etc. | 简单易用的语音工具箱
An Open-Sourced LLM-empowered Foundation TTS System
Turn an epub or text file into an audiobook
NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment
an open-source implementation of sequence-to-sequence based speech processing engine
A Non-Autoregressive Text-to-Speech (NAR-TTS) framework, including official PyTorch implementation of PortaSpeech (NeurIPS 2021) and DiffSpeech (AAAI 2022)
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.
GLM-TTS: Controllable & Emotion-Expressive Zero-shot TTS with Multi-Reward Reinforcement Learning
免费的在线文本转语音API
A simple VITS HTTP API, developed by extending Moegoe with additional features.
With one command, create a natural-sounding audiobook from a variety of input formats (epub, mobi, txt, PDF, HTML and more!)
YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyone
A TensorFlow Implementation of DC-TTS: yet another text-to-speech model
Free and open source text-to-speech software
Chinese text-to-speech engine
Best practice TTS based on BERT and VITS with some Natural Speech Features Of Microsoft; Support ONNX streaming out!
StreamSpeech is an “All in One” seamless model for offline and simultaneous speech recognition, speech translation and speech synthesis.
实时交互数字人,可自定义形象与音色,支持音色克隆,对话延迟低至3s。Real-time voice interactive digital human, customizable appearance and voice, supporting voice cloning, with initial package delay as low as 3s.
24,523 repositories in the index in total.