alan-ai/alan-sdk-android
quality grade D, 37 out of 100The Self-Coding System for Your App — Alan AI SDK for Android
- stars
- 1.8k
- stars gained this week
- —this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Speech recognition, text-to-speech, voice cloning, music generation and audio processing.
Signals: speech-recognition, text-to-speech, tts, stt, asr, whisper, voice-cloning, speech-synthesis
915 results
The Self-Coding System for Your App — Alan AI SDK for Android
The Self-Coding System for Your App — Alan AI SDK for iOS
Offline voice assistant that respects your privacy. Forged in Rust. WIP.
🧠 Leon is your open-source personal assistant.
a comfyui custom node for GPT-SoVITS! you can voice cloning and tts in comfyui now
A ComfyUI custom node suite for Qwen3-TTS, supporting 1.7B and 0.6B models, Custom Voice, Voice Design, Voice Cloning and Fine-Tuning.
Your faithful, impartial partner for audio evaluation — know yourself, know your rivals. 真实评测,知己知彼。A unified benchmark framework for ASR/TTS/Audio Codec/audio LLM evaluation
Multilingual TTS model with voice cloning and duration control, based on T5Gemma encoder-decoder LLM
An app for creating audio-based content such as song covers and speech using Retrieval-based Voice Conversion.
A program to dub non-english media with modern AI speech synthesis, diarization, and voice cloning!
🎭 AI Avatar / digital human platform — upload a photo, clone a voice, talk to any face in real time with lip-sync video. Open-source, self-hosted. Claude · Whisper · Chatterbox · MuseTalk.
Sesame CSM 1B Voice Cloning
Voice AI runtime. Local first transcription, speaker diarization, TTS, and voice cloning with an OpenAI compatible API.
Self-host the powerful Dia TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), support for SafeTensors/BF16, voice cloning, dialogue generation, and GPU/CPU execution.
EaseVoice Trainer is a simple and user-friendly voice cloning and speech model trainer.
Tacotron 2 - PyTorch implementation with faster-than-realtime inference modified to enable cross lingual voice cloning.
VoxNovel: generate audiobooks giving each character a different voice actor.
Talk to 峰哥 — 克隆任何人的声音和性格,实时语音对话,工程延迟 < 1 秒 | Clone anyone's voice & personality for real-time conversation. < 1s engineering latency.
List of open-source TTS, voice cloning, and music generation models
A User Interface for XTTS-2 Text-Based Voice Cloning using only 10 seconds of speech
Fuse ChatTTS with OpenVoice, upload a 10-second audio clip, and clone your personalized ChatTTS voice.
ComfyUI node for highly expressive speech and realistic zero-shot voice cloning
OmniVoice TTS nodes for ComfyUI - Zero-shot multilingual text-to-speech with voice cloning, voice design, and multi-speaker dialogue
Run Qwen3-TTS text-to-speech locally on Mac (M1/M2/M3/M4). Voice cloning, voice design, custom voices. 100% offline using MLX.
24,523 repositories in the index in total.