codeforequity-at/botium-speech-processing
quality grade B, 74 out of 100Botium Speech Processing
- stars
- 943
- stars gained this week
- —this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Speech recognition, text-to-speech, voice cloning, music generation and audio processing.
Signals: speech-recognition, text-to-speech, tts, stt, asr, whisper, voice-cloning, speech-synthesis
913 results
Botium Speech Processing
TTS model capable of streaming conversational audio in realtime.
[ICLR 2025] SOTA discrete acoustic codec models with 40/75 tokens per second for audio language modeling
This is now the official location of the Merlin project.
the open-source virtual assistant for Ubuntu based Linux distributions
An awesome browser extension that reads aloud webpage content with one click
DeepMind's Tacotron-2 Tensorflow implementation
Offline Text To Speech synthesis for python
🚀 一键部署(含离线整合包)!基于 ChatTTS ,支持流式输出、音色抽卡、长音频生成和分角色朗读。简单易用,无需复杂安装。
Converts text to speech in realtime
The python library for real-time communication
eSpeak NG is an open source speech synthesizer that supports more than hundred languages and accents.
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.
A TTS model capable of generating ultra-realistic dialogue in one pass.
MusicTransformer written for MaestroV2 using the Pytorch framework for music generation
End-to-End Automatic Speech Recognition on PyTorch
Midi event transformer for symbolic music generation
Implementation of MusicLM, a text to music model published by Google Research, with a few modifications.
Open-Source Toolkit for End-to-End Korean Automatic Speech Recognition leveraging PyTorch and Hydra.
A PyTorch implementation of Speech Transformer, an End-to-End ASR with Transformer network on Mandarin Chinese.
[Unofficial] PyTorch implementation of "Conformer: Convolution-augmented Transformer for Speech Recognition" (INTERSPEECH 2020)
Soprano-Factory: Train your own 2000x realtime text-to-speech model
10000 chatTTS voices !chatTTS 音色库,再也不为音色抽卡烦恼啦。这是我第一个项目,熬夜龟速生产10000条音色并上传Github,给点鼓励呗哈!主域名:https://www.TTSlist.com 备用:http://ttslist.aiqbh.com/
Official implementation of Meta-StyleSpeech and StyleSpeech
24,524 repositories in the index in total.