kxxt/aspeak
quality grade C, 57 out of 100A simple text-to-speech client for Azure TTS API.
- stars
- 496
- stars gained this week
- —this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Speech recognition, text-to-speech, voice cloning, music generation and audio processing.
Signals: speech-recognition, text-to-speech, tts, stt, asr, whisper, voice-cloning, speech-synthesis
910 results
A simple text-to-speech client for Azure TTS API.
Mimic Recording Studio is a Docker-based application you can install to record voice samples, which can then be trained into a TTS voice with Mimic2
KAN-TTS is a speech-synthesis training framework, please try the demos we have posted at https://modelscope.cn/models?page=1&tasks=text-to-speech
Automatic Speech Recognition(ASR), Text-To-Speech(TTS) engine. 中英语音识别、多角色语音合成,支持多语言,准确率高
Android speech recognition and text to speech made easy
基于Bert-VITS2做的表情、动画测试. Animation testing based on Bert-VITS2.
SummerTTS 是一个基于C++的独立编译的中文和英文语音合成项目,可以本地运行不需要网络,而且没有额外的依赖,一键编译完成即可用于中文和英文的语音合成。SummerTTS is a standalone Chinese and English speech synthesis(TTS) project that has almost no dependency and could be easily used for Chinese TTS with just one key build out
Audio samples accompanying publications related to Tacotron, an end-to-end speech synthesis model.
Run Orpheus 3B Locally With LM Studio
unofficial vits2-TTS implementation in pytorch
A simple FastAPI Server to run XTTSv2
Simple Python script to interact with the TikTok TTS API
A Fast TTS Engine
VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture Design
Implementation of F5-TTS in MLX
A series of 3 programs that will automatically receive scripts from Reddit, allow the user to edit them, then be sent off to a video generator where they will be uploaded to YouTube automatically.
🔊 Kokoro Web: Free AI text-to-speech, online or self-hosted, OpenAI compatible!
A Pytorch Implementation of "Neural Speech Synthesis with Transformer Network"
Easily create Piper text-to-speech models in any voice. Make a text-to-speech model with your own voice recordings, or use thousands of RVC voices. Works offline on a Raspberry pi. Rapidly record custom datasets for any metadata.csv file and listen to your model as it is training.
A Generative Flow for Text-to-Speech via Monotonic Alignment Search
Thorsten-Voice: A free to use, offline working, high quality german TTS voice should be available for every project without any license struggling.
🐸 collection of TTS papers
MLX Omni Server is a local inference server powered by Apple's MLX framework, specifically designed for Apple Silicon (M-series) chips. It implements OpenAI-compatible API endpoints, enabling seamless integration with existing OpenAI SDK clients while leveraging the power of local ML inference.
24,535 repositories in the index in total.