Trending repositories: speech-to-text

9 tracked repositories tagged with speech-to-text, ordered by stars. Use the topic filters below to narrow further.

Filter by topic

9 of 9 repositories

  • cjpais/Handy

    A free, open source, and extensible speech-to-text application that works completely offline.

    AI summary: A privacy-focused, cross-platform desktop application for completely offline speech-to-text transcription.

    28,924productivityRustMIT
  • Zackriya-Solutions/meetily

    Privacy first, AI meeting assistant with 4x faster Parakeet/Whisper live transcription, speaker diarization, and Ollama summarization built on Rust. 100% local processing. no cloud required. Meetily (Meetly Ai - https://meetily.ai) is the #1 Self-hosted, Open-source Ai meeting note taker for macOS & Windows. Understand How to write meeting minutes

    AI summary: An open-source meeting scheduling application prioritizing privacy and self-hosting.

    28,405productivityRustMIT
  • Open-LLM-VTuber/Open-LLM-VTuber

    Talk to any LLM with hands-free voice interaction, voice interruption, and Live2D taking face running locally across platforms

    AI summary: A complete stack for running AI VTubers locally with voice, Live2D, and interruption support.

    13,130ai-mlPythonOther
  • abus-aikorea/voice-pro

    Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

    AI summary: Comprehensive Gradio WebUI for advanced audio processing, featuring zero-shot voice cloning, transcription, and dubbing.

    12,072ai-mlPythonGPL-3.0
  • huggingface/speech-to-speech

    Build local voice agents with open-source models

    AI summary: A modular, open-source voice pipeline for building real-time, low-latency conversational AI agents.

    11,550ai-mlPythonApache-2.0
  • moonshine-ai/moonshine

    Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces

    AI summary: A fast, cross-platform family of open speech-to-text models designed specifically for low-latency live voice interfaces.

    10,659ai-mlC++Other
  • Blaizzy/mlx-audio

    A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.

    AI summary: High-performance audio processing library optimized for Apple Silicon via the MLX framework.

    7,688ai-mlPythonMIT
  • matthartman/ghost-pepper

    100% private on-device voice models for speech-to-text and meeting transcription on macOS

    AI summary: A 100% private, on-device voice model application for speech-to-text and transcription on macOS.

    3,070productivitySwift
  • zachlatta/freeflow

    Free & fast alternative to Wispr Flow

    AI summary: A free, fast, and open-source Mac dictation app alternative to Wispr Flow.

    2,395productivitySwiftMIT