isaiahbjork/orpheus-tts-local
quality grade D, 38 out of 100Run Orpheus 3B Locally With LM Studio
- stars
- 546
- stars gained this week
- —this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Speech recognition, text-to-speech, voice cloning, music generation and audio processing.
Signals: speech-recognition, text-to-speech, tts, stt, asr, whisper, voice-cloning, speech-synthesis
915 results
Run Orpheus 3B Locally With LM Studio
unofficial vits2-TTS implementation in pytorch
「来剪」轻量级视频编辑器。网页版、桌面版等均可免费使用,功能灵感源自 CapCut 等编辑器。A Lightweight Video Editor. Free for the web, desktop, and more, with features inspired by editors like CapCut.
A simple FastAPI Server to run XTTSv2
Simple Python script to interact with the TikTok TTS API
A Fast TTS Engine
VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture Design
Implementation of F5-TTS in MLX
A series of 3 programs that will automatically receive scripts from Reddit, allow the user to edit them, then be sent off to a video generator where they will be uploaded to YouTube automatically.
🔊 Kokoro Web: Free AI text-to-speech, online or self-hosted, OpenAI compatible!
A Pytorch Implementation of "Neural Speech Synthesis with Transformer Network"
Easily create Piper text-to-speech models in any voice. Make a text-to-speech model with your own voice recordings, or use thousands of RVC voices. Works offline on a Raspberry pi. Rapidly record custom datasets for any metadata.csv file and listen to your model as it is training.
Confucius4-TTS: a Multilingual and Cross-Lingual Zero-Shot TTS Engine
A Generative Flow for Text-to-Speech via Monotonic Alignment Search
Thorsten-Voice: A free to use, offline working, high quality german TTS voice should be available for every project without any license struggling.
🐸 collection of TTS papers
MLX Omni Server is a local inference server powered by Apple's MLX framework, specifically designed for Apple Silicon (M-series) chips. It implements OpenAI-compatible API endpoints, enabling seamless integration with existing OpenAI SDK clients while leveraging the power of local ML inference.
Flutter Text to Speech package
Deep learning for audio processing
AI VTuber with LLM, ASR, TTS, OCR, CV and more technologies to live stream or play Minecraft with you.
Speech to Text to Speech. Song now playing. Sends text as OSC messages to VRChat to display on avatar. (STTTS) (Speech to TTS) (VRC STT System) (VTuber TTS)
🎤 Lobe TTS - A high-quality & reliable TTS/STT library for Server and Browser
Multi-source Translation
24,523 repositories in the index in total.