Skip to main content
buildradar
Sign in
Topic · speech-to-text

speech-to-text

Tracked open-source repos tagged speech-to-text, sorted by stars.

152 repos
  • typewhisper-mac@TypeWhisper

    Local speech-to-text for macOS on-device AI, fully private, optional cloud

    1,759+0Star change over the last 7 days
  • react-native-executorch@software-mansion

    Declarative way to run AI models in React Native on device, powered by ExecuTorch.

    1,707+0Star change over the last 7 days
  • 收集户晨风的所有内容

    1,650+0Star change over the last 7 days
  • dsnote@mkiol

    Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.

    1,622+0Star change over the last 7 days
  • OBS plugin for local speech recognition and captioning using AI

    1,597+0Star change over the last 7 days
  • vllm-mlx@waybarrios

    High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon. Native MLX, continuous batching, multimodal models, MCP tool calling, and Claude Code support.

    1,557+0Star change over the last 7 days
  • TikTokLive@isaackogan

    TikTok LIVE API for Python: The definitive 3rd-party library to receive livestream events (comments, gifts, etc.) in realtime from TikTok LIVE.

    1,544+0Star change over the last 7 days
  • RCLI@RunanywhereAI

    Talk to your Mac, query your docs, no cloud required. On-device voice AI + RAG

    1,542+0Star change over the last 7 days
  • Custom nodes that extend the capabilities of Comfyui

    1,525+0Star change over the last 7 days
  • amical@amicalhq

    🎙️ AI Dictation App - Open Source and Local-first ⚡ Type 3x faster, no keyboard needed. 🆓 Powered by open source models, works offline, fast and accurate.

    1,520+0Star change over the last 7 days
  • Fun-ASR@QwenAudio

    Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR, vLLM, streaming, and llama.cpp runtimes.

    1,512+0Star change over the last 7 days
  • minutes@silverstein

    Open-source, local-first Granola/Otter alternative that Claude Code, Codex, Cursor, and any MCP client can query. Meetings, calls, and voice memos transcribed on-device into markdown you own.

    1,462+0Star change over the last 7 days
  • SoniTranslate@R3gm

    Synchronized Translation for Videos. Video dubbing

    1,413+0Star change over the last 7 days
  • 💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies

    1,399+0Star change over the last 7 days
  • mlx-tune@ARahim3

    Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unsloth-compatible API.

    1,398+0Star change over the last 7 days
  • Multi-backend whisper app. Blazing fast. Mac-arm optimized. Easy install. Input a local file or url and this service will transcribe it using Whisper AI. Completely private and Free 🤯🤯🤯

    1,394+0Star change over the last 7 days
  • whisper-ctranslate2@Softcatala

    Whisper command line client compatible with original OpenAI client based on CTranslate2.

    1,346+0Star change over the last 7 days
  • voxtype@peteonrails

    Voice-to-text with push-to-talk for Wayland compositors

    1,338+0Star change over the last 7 days
  • gp.nvim@Robitx

    Gp.nvim (GPT prompt) Neovim AI plugin: ChatGPT sessions & Instructable text/code operations & Speech to text [OpenAI, Ollama, Anthropic, ..]

    1,320+0Star change over the last 7 days
  • airunner@Capsize-Games

    Offline inference engine for art, real-time voice conversations, LLM powered chatbots and automated workflows

    1,314+0Star change over the last 7 days
  • StreamSpeech@ictnlp

    StreamSpeech is an “All in One” seamless model for offline and simultaneous speech recognition, speech translation and speech synthesis.

    1,288+0Star change over the last 7 days
  • quillman@modal-labs

    A voice chat app

    1,212+0Star change over the last 7 days
  • hyprwhspr@goodroot

    Native speech-to-text for Linux - Fast, accurate, private, and hackable system-wide dictation

    1,189+0Star change over the last 7 days
  • Uncensored-Local-Studio@techjarves

    Uncensored local AI studio for Windows, Linux, and macOS. Zero-setup GUI for Image Generation, GGUF LLMs, Text to Speech & Speech to Text

    1,118+13Star change over the last 7 days
  • AI-Waifu-Vtuber@ardha27

    AI Vtuber for Streaming on Youtube/Twitch

    1,116+3Star change over the last 7 days
  • muesli@Muesli-HQ

    Muesli: agent-native local meeting transcription + dictation for macOS (Granola + WisprFlow alternative)

    1,105+35Star change over the last 7 days
  • Whisperboard@Saik0s

    The open-source iOS app that's making quality voice transcription more accessible on mobile devices.

    1,102+6Star change over the last 7 days
  • whisper-writer@savbell

    💬📝 A small dictation app using OpenAI's Whisper speech recognition model.

    1,100+0Star change over the last 7 days
  • murmure@Kieirra

    Fully local, private and cross platform Speech-to-Text with LLM Post-processing

    1,056+9Star change over the last 7 days
  • Patter@PatterAI

    Open-source voice-AI SDK. The Vapi/Retell alternative for builders who want to own the stack. Give your AI agent a phone number in 4 lines — Python and TypeScript, MIT licensed, Twilio, Telnyx, and Plivo.

    1,051+7Star change over the last 7 days
← Back to topics