LibreSpark/LibreTTS
quality grade C, 63 out of 100TTS-文本转语音/文本转语音前端,兼容OpenAI、EdgeTTS等接口
- stars
- 383
- stars gained this week
- +1this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Speech recognition, text-to-speech, voice cloning, music generation and audio processing.
Signals: speech-recognition, text-to-speech, tts, stt, asr, whisper, voice-cloning, speech-synthesis
909 results
TTS-文本转语音/文本转语音前端,兼容OpenAI、EdgeTTS等接口
Example projects built with the Hume AI APIs
Open Voice Operating System - Buildroot edition is a minimalistic linux OS bringing the OVOS voice assistant to embbeded, low-spec headless and/or small (touch)screen devices.
Generate TikTok-style captions with Whisper.cpp
A lightweight Python package for Automatic Speech Recognition using ONNX models
Transcribe and translate voice into LRC file using Whisper and LLMs (GPT, Claude, et,al). 使用whisper和LLM(GPT,Claude等)来转录、翻译你的音频为字幕文件。
🎩 An Alfred 5 Workflow for using OpenAI Chat API to interact with GPT models 🤖💬 It also allows image generation/editing/understanding 🖼️, speech-to-text conversion 🎤, and text-to-speech synthesis 🔈
This is a TensorFlow implementation of the WaveNet generative neural network architecture https://deepmind.com/blog/wavenet-generative-model-raw-audio/ for text generation.
Sound event localization, detection, and tracking of multiple overlapping and moving sources in 2D spherical space using convolutional recurrent neural network
Recurrent Neural Network for generating piano MIDI-files from audio (MP3, WAV, etc.)
Tensorflow 2.0 implementation of the paper: A Fully Convolutional Neural Network for Speech Enhancement
Recurrent neural network for audio noise reduction
Trained neural networks and requisite information and data for rnnoise-nu
Deep neural networks for getting text-independent speaker embedding written in TensorFlow
Predicting depression from acoustic features of speech using a Convolutional Neural Network.
Real-time Voice Activity Detection in Noisy Eniviroments using Deep Neural Networks
Pitch Estimating Neural Networks (PENN)
Automatic Music Transcription with Deep Neural Networks
This is the code for "Neural Network Voices" by Siraj Raval on Youtube
A realtime live transcription and translation app built with Huggingface Transformer.js and Supabase Realtime.
No description
No description
plug whisper audio transcription to a local ollama server and ouput tts audio responses
Voice interface for Claude Code via SIP/3CX - Call your AI, and your AI can call you
24,538 repositories in the index in total.