SakiRinn/LiveCaptions-Translator
quality grade B, 79 out of 100Lightweight and powerful real-time audio/speech translation tool based on Windows LiveCaptions.
- stars
- 3.4k
- stars gained this week
- —this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Speech recognition, text-to-speech, voice cloning, music generation and audio processing.
Signals: speech-recognition, text-to-speech, tts, stt, asr, whisper, voice-cloning, speech-synthesis
913 results
Lightweight and powerful real-time audio/speech translation tool based on Windows LiveCaptions.
A robust, efficient, low-latency speech-to-text library with advanced voice activity detection, wake word activation and instant transcription.
A PyTorch-based Speech Toolkit
kaldi-asr/kaldi is the official location of the Kaldi project.
Real time web based Speech-to-Text app with Streamlit
A Streamilt web app for music source separation & karaoke
Get started using Deepgram's Live Transcription with this Next.js demo app
T-one is a high-performance streaming ASR pipeline for Russian, specialized for the telephony domain.
Open source speech to text models for Indic Languages
On-device speech-to-text engine powered by deep learning
A speech recognition library running in the browser thanks to a WebAssembly build of Vosk
A React component to make correcting automated transcriptions of audio and video easier and faster. By BBC News Labs. - Work in progress
:speech_balloon: /so.nus/ STT (speech to text) for Node with offline hotword detection
Voice Recognition to Text Tool / 一个离线运行的本地音视频转字幕工具,输出json、srt字幕、纯文字格式
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
Machine Learning Training Utilities (for TensorFlow and PyTorch)
End-to-End speech recognition implementation base on TensorFlow (CTC, Attention, and MTL training)
State-of-the-art (ranked #1 Aug 2022) German Speech Recognition in 284 lines of C++. This is a 100% private 100% offline 100% free CLI tool.
The official repository of the Eesen project
中文语音识别
:zap: TensorFlowASR: Almost State-of-the-art Automatic Speech Recognition in Tensorflow 2. Supported languages that can use characters or subwords
An AI for Music Generation
🎙Speech recognition using the tensorflow deep learning framework, sequence-to-sequence neural networks
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
24,524 repositories in the index in total.