algolia/voice-overlay-ios
quality grade B, 65 out of 100🗣 An overlay that gets your user’s voice permission and input as text in a customizable UI
- stars
- 557
- stars gained this week
- —this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Speech recognition, text-to-speech, voice cloning, music generation and audio processing.
Signals: speech-recognition, text-to-speech, tts, stt, asr, whisper, voice-cloning, speech-synthesis
910 results
🗣 An overlay that gets your user’s voice permission and input as text in a customizable UI
The Self-Coding System for Your App — Alan AI SDK for React Native
Open Source Voice Agent Platform
Main repository of Project Alice, contains main unit source code
A python based desktop voice assistant capable of executing system-level commands, integrating speech recognition and text-to-speech, and handling asynchronous user interactions.
This app can now use Android, just like a human.
Python AI assistant 🧠
The Self-Coding System for Your App — Alan AI SDK for Cordova
Ирина - русский голосовой ассистент для работы оффлайн. Поддерживает скиллы через плагины.
The Self-Coding System for Your App — Alan AI SDK for Ionic
百聆 是一个类似GPT-4o的语音对话机器人,通过ASR+LLM+TTS实现,集成DeepSeek R1等优秀大模型,接入openClaw,真正的个人语音助手,时延低至800ms,Mac等低配置也可运行,支持打断
The Self-Coding System for Your App — Alan AI SDK for Flutter
The Self-Coding System for Your App — Alan AI SDK for Android
The Self-Coding System for Your App — Alan AI SDK for iOS
Offline voice assistant that respects your privacy. Forged in Rust. WIP.
a comfyui custom node for GPT-SoVITS! you can voice cloning and tts in comfyui now
A ComfyUI custom node suite for Qwen3-TTS, supporting 1.7B and 0.6B models, Custom Voice, Voice Design, Voice Cloning and Fine-Tuning.
Your faithful, impartial partner for audio evaluation — know yourself, know your rivals. 真实评测,知己知彼。A unified benchmark framework for ASR/TTS/Audio Codec/audio LLM evaluation
Multilingual TTS model with voice cloning and duration control, based on T5Gemma encoder-decoder LLM
An app for creating audio-based content such as song covers and speech using Retrieval-based Voice Conversion.
A program to dub non-english media with modern AI speech synthesis, diarization, and voice cloning!
Sesame CSM 1B Voice Cloning
Voice AI runtime. Local first transcription, speaker diarization, TTS, and voice cloning with an OpenAI compatible API.
Self-host the powerful Dia TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), support for SafeTensors/BF16, voice cloning, dialogue generation, and GPU/CPU execution.
24,535 repositories in the index in total.