coqui-ai/TTS
quality grade C, 55 out of 100🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
- stars
- 46k
- stars gained this week
- +26this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Speech recognition, text-to-speech, voice cloning, music generation and audio processing.
Signals: speech-recognition, text-to-speech, tts, stt, asr, whisper, voice-cloning, speech-synthesis
910 results
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
Clone a voice in 5 seconds to generate arbitrary speech in real-time
A real-time speech-to-speech chatbot powered by Whisper Small, Llama 3.2, and Kokoro-82M.
This repository will guide you to create your own Smart Virtual Assistant like Google Assistant using Open AI's ChatGPT, Whisper. The entire solution is created using Python & Gradio.
The main repo for Stage Whisper — a free, secure, and easy-to-use transcription app for journalists, powered by OpenAI's Whisper automatic speech recognition (ASR) machine learning models.
ARIA - AI Realtime Intelligent Audio | Universal real-time AI subtitles for Windows
Transcribe is a real time transcription, conversation, Language learning platform. It provides live transcripts from microphone and speaker. It generates a suggested conversation response using OpenAI's GPT API. It will read out the responses, simulating a real live conversation in English or another language.
Input0 — A macOS voice input tool: hold a hotkey to record, release to transcribe locally via STT, refine with LLM, and auto-paste into the active text field.
AI Device Template Featuring Whisper, TTS, Groq, Llama3, OpenAI and more
Node.js bindings for OpenAI's Whisper. (C++ CPU version by ggerganov)
AI-powered tool for real-time interview question transcription and response generation.
Simple self-hosted web application, which can be used to convert audio to subtitles by OpenAI's Whisper model
From AI tools to TikTok video creation using FFMPEG, Microsoft Edge read aloud and OpenAI Whisper model
Pybind11 bindings for Whisper.cpp
Experimental code: sound file preprocessing to optimize Whisper transcriptions without hallucinated texts
An API to transcribe audio with OpenAI's Whisper Large v3!
AIUI is a platform enabling seamless two-way verbal communication with AI.
A nearly-live implementation of OpenAI's Whisper, using sounddevice. Requires existing Whisper install.
Fine-tune and evaluate Whisper models for Automatic Speech Recognition (ASR) on custom datasets or datasets from huggingface.
Speech-to-text in Obsidian using Whisper
🎬 Auto Captions for Final Cut Pro Powered by OpenAI's Whisper Model
Free on-device web app for audio transcribing and rendering subtitles
Your personal voice interface for any app. Speak naturally and your words appear wherever your cursor is, with fully customizable AI voice dictation. Open source alternative to Wispr Flow.
How to use OpenAIs Whisper to transcribe and diarize audio files
24,535 repositories in the index in total.