Trending repositories: speech-to-text
9 tracked repositories tagged with speech-to-text, ordered by stars. Use the topic filters below to narrow further.
9 of 9 repositories
cjpais/Handy
A free, open source, and extensible speech-to-text application that works completely offline.
AI summary: A privacy-focused, cross-platform desktop application for completely offline speech-to-text transcription.
28,924productivityRustMITZackriya-Solutions/meetily
Privacy first, AI meeting assistant with 4x faster Parakeet/Whisper live transcription, speaker diarization, and Ollama summarization built on Rust. 100% local processing. no cloud required. Meetily (Meetly Ai - https://meetily.ai) is the #1 Self-hosted, Open-source Ai meeting note taker for macOS & Windows. Understand How to write meeting minutes
AI summary: An open-source meeting scheduling application prioritizing privacy and self-hosting.
28,405productivityRustMITOpen-LLM-VTuber/Open-LLM-VTuber
Talk to any LLM with hands-free voice interaction, voice interruption, and Live2D taking face running locally across platforms
AI summary: A complete stack for running AI VTubers locally with voice, Live2D, and interruption support.
13,130ai-mlPythonOtherabus-aikorea/voice-pro
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.
AI summary: Comprehensive Gradio WebUI for advanced audio processing, featuring zero-shot voice cloning, transcription, and dubbing.
12,072ai-mlPythonGPL-3.0huggingface/speech-to-speech
Build local voice agents with open-source models
AI summary: A modular, open-source voice pipeline for building real-time, low-latency conversational AI agents.
11,550ai-mlPythonApache-2.0moonshine-ai/moonshine
Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces
AI summary: A fast, cross-platform family of open speech-to-text models designed specifically for low-latency live voice interfaces.
10,659ai-mlC++OtherBlaizzy/mlx-audio
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
AI summary: High-performance audio processing library optimized for Apple Silicon via the MLX framework.
7,688ai-mlPythonMITmatthartman/ghost-pepper
100% private on-device voice models for speech-to-text and meeting transcription on macOS
AI summary: A 100% private, on-device voice model application for speech-to-text and transcription on macOS.
3,070productivitySwiftzachlatta/freeflow
Free & fast alternative to Wispr Flow
AI summary: A free, fast, and open-source Mac dictation app alternative to Wispr Flow.
2,395productivitySwiftMIT