lenML/Speech-AI-Forge
quality grade D, 49 out of 100🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.
- stars
- 1.4k
- stars gained this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Speech recognition, text-to-speech, voice cloning, music generation and audio processing.
Signals: speech-recognition, text-to-speech, tts, stt, asr, whisper, voice-cloning, speech-synthesis
910 results
🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.
Using OpenAI's Whisper to automatically generate YouTube subtitles
🎞️ Subtitles generation tool (Web-UI + CLI + Python package) powered by OpenAI's Whisper and its variants 🎞️
Cross-Platform, GPU Accelerated Whisper 🏎️
Automated YouTube Shorts pipeline: news → script → AI visuals → voiceover → captions → upload
快速提取音视频内容,整理成一份结构化的markdown笔记
Automatically generate and overlay subtitles for any video.
Multilingual Automatic Speech Recognition with word-level timestamps and confidence
A Web UI for easy subtitle using whisper model.
Transcribe any audio to text, translate and edit subtitles 100% locally with a web UI. Powered by whisper models!
Whisper & Faster-Whisper standalone executables for those who don't want to bother with Python.
OpenAI API + Ruby! 🤖❤️ GPT-5 & Realtime WebRTC compatible!
ML-powered speech recognition directly in your browser
ChatGPT Java SDK支持流式输出、Gpt插件、联网。支持OpenAI官方所有接口。ChatGPT的Java客户端。OpenAI GPT-3.5-Turb GPT-4 Api Client for Java
🤖 A Telegram bot that integrates with OpenAI's official ChatGPT APIs to provide answers, written in Python
Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.
Mac app for crushing tech interviews with AI
JAX implementation of OpenAI's Whisper model for up to 70x speed-up on TPU.
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.
🎩 An Alfred 5 Workflow for using OpenAI Chat API to interact with GPT models 🤖💬 It also allows image generation/editing/understanding 🖼️, speech-to-text conversion 🎤, and text-to-speech synthesis 🔈
Amica is an open source interface for interactive communication with 3D characters with voice synthesis and speech recognition.
24,535 repositories in the index in total.