gradio-app/fastrtc
quality grade C, 59 out of 100The python library for real-time communication
- stars
- 4.6k
- stars gained this week
- +2this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Speech recognition, text-to-speech, voice cloning, music generation and audio processing.
Signals: speech-recognition, text-to-speech, tts, stt, asr, whisper, voice-cloning, speech-synthesis
910 results
The python library for real-time communication
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.
A TTS model capable of generating ultra-realistic dialogue in one pass.
MusicTransformer written for MaestroV2 using the Pytorch framework for music generation
End-to-End Automatic Speech Recognition on PyTorch
Midi event transformer for symbolic music generation
Implementation of MusicLM, a text to music model published by Google Research, with a few modifications.
Open-Source Toolkit for End-to-End Korean Automatic Speech Recognition leveraging PyTorch and Hydra.
A PyTorch implementation of Speech Transformer, an End-to-End ASR with Transformer network on Mandarin Chinese.
[Unofficial] PyTorch implementation of "Conformer: Convolution-augmented Transformer for Speech Recognition" (INTERSPEECH 2020)
Soprano-Factory: Train your own 2000x realtime text-to-speech model
10000 chatTTS voices !chatTTS 音色库,再也不为音色抽卡烦恼啦。这是我第一个项目,熬夜龟速生产10000条音色并上传Github,给点鼓励呗哈!主域名:https://www.TTSlist.com 备用:http://ttslist.aiqbh.com/
Official implementation of Meta-StyleSpeech and StyleSpeech
HiFTNet: A Fast High-Quality Neural Vocoder with Harmonic-plus-Noise Filter and Inverse Short Time Fourier Transform
Official repository of DailyTalk: Spoken Dialogue Dataset for Conversational Text-to-Speech, ICASSP 2023
🐤 Nix-TTS: Lightweight and End-to-end Text-to-Speech via Module-wise Distillation
🔊 Cross browser Speech Synthesis also known as Text to speech or TTS; no dependencies; uses Web Speech API
Praises is a text-to-speech tool that can help you read text easily.
ttslearn: Library for Pythonで学ぶ音声合成 (Text-to-speech with Python)
Phoneme-Level BERT for Enhanced Prosody of Text-to-Speech with Grapheme Predictions
Self-host the ultra-lightweight Kitten TTS model with this enhanced API server with an intuitive Web UI, large text processing for audiobooks, and GPU acceleration.
Kokoro TTS for iOS and macOSX
Multilingual Inexpensive Therapeutic Sophisticated Ultra-responsive Holographic Agent. In simple terms, an AI you can talk to and it'll talk back with a body using VTube Studio.
A unified interface for multiple Text-to-Speech (TTS) providers.
24,537 repositories in the index in total.