Skip to main content
buildradar
Sign in
Topic · speech-to-text

speech-to-text

Tracked open-source repos tagged speech-to-text, sorted by stars.

152 repos
  • LiveStream-Agent-Studio@HanyuanWang

    面向抖音直播电商的 Windows 本地 AI Agent Studio,贯通主播发现、直播洞察、直播复盘与短视频内容编导的统一智能工作流。

    1,018+424Star change over the last 7 days
  • TensorFlowASR@TensorSpeech

    :zap: TensorFlowASR: Almost State-of-the-art Automatic Speech Recognition in Tensorflow 2. Supported languages that can use characters or subwords

    1,010+0Star change over the last 7 days
  • voquill@voquill

    Open source voice dictation technology

    1,006-1Star change over the last 7 days
  • whisper-flow@dimastatz

    Whisper-Flow is a framework designed to enable real-time transcription of audio content using OpenAI’s Whisper model. Rather than processing entire files after upload (“batch mode”), Whisper-Flow accepts a continuous stream of audio chunks and produces incremental transcripts immediately.

    964+25Star change over the last 7 days
  • VoiceStreamAI@alesaccoia

    Near-Realtime audio transcription using self-hosted Whisper and WebSocket in Python/JS

    960+0Star change over the last 7 days
  • botium-speech-processing@codeforequity-at

    Botium Speech Processing

    943+0Star change over the last 7 days
  • whisper.net@sandrohanea

    Whisper.net. Speech to text made simple using Whisper Models

    942+1Star change over the last 7 days
  • jiwer@jitsi

    Evaluate your speech-to-text system with similarity measures such as word error rate (WER)

    928+2Star change over the last 7 days
  • MauiSamples@VladislavAntonyuk

    .NET MAUI Samples

    915+0Star change over the last 7 days
  • voicy@backmeupplz

    @voicybot Telegram bot main repository

    911+2Star change over the last 7 days
  • TheWhisper@TheStageAI

    Optimized Whisper models for streaming and on-device use

    897-1Star change over the last 7 days
  • TypeNo@marswaveai

    A free, open source, privacy-first voice input app for macOS.

    896+4Star change over the last 7 days
  • PPASR@yeyupiaoling

    基于PaddlePaddle实现端到端中文语音识别,从入门到实战,超简单的入门案例,超实用的企业项目。支持当前最流行的DeepSpeech2、Conformer、Squeezeformer模型

    870-1Star change over the last 7 days
  • react-speech-recognition@JamesBrill

    💬Speech recognition for your React app

    842+0Star change over the last 7 days
  • whisper-playground@saharmor

    Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/

    835+0Star change over the last 7 days
  • kur@deepgram

    Descriptive Deep Learning

    823-1Star change over the last 7 days
  • MioSub@corvo007

    一站式全自动字幕生成软件,下载、转录、翻译、压制全流程覆盖,无需人工介入 / One-stop automated subtitle generator. Handles downloading, transcription, translation, and hardcoding—zero human intervention required.

    811+3Star change over the last 7 days
  • vocalinux@VocaHQ

    Free, open-source, 100% offline voice dictation for Linux. Speak and type anywhere via whisper.cpp, Whisper & VOSK engines, GPU-accelerated, works on X11 + Wayland!

    809+17Star change over the last 7 days
  • lobe-tts@lobehub

    🎤 Lobe TTS - A high-quality & reliable TTS/STT library for Server and Browser

    805-1Star change over the last 7 days
  • Speech to Text to Speech. Song now playing. Sends text as OSC messages to VRChat to display on avatar. (STTTS) (Speech to TTS) (VRC STT System) (VTuber TTS)

    804+5Star change over the last 7 days
  • june@mezbaul-h

    Local voice chatbot for engaging conversations, powered by Ollama, Hugging Face Transformers, and Coqui TTS Toolkit

    788+1Star change over the last 7 days
  • whisper_mic@mallorbc

    Project that allows one to use a microphone with OpenAI whisper.

    787+0Star change over the last 7 days
  • SwiftWhisper@exPHAT

    🎤 The easiest way to transcribe audio in Swift

    785+0Star change over the last 7 days
  • NotelyVoice@Notely-Voice

    A 100% private AI voice transcription app that converts speech to text in 100+ languages. Built with Compose Multiplatform for Android & iOS using Whisper AI - no cloud uploads, all processing happens on-device for complete privacy.

    779+4Star change over the last 7 days
  • mlx-audio-swift@Blaizzy

    A modular Swift SDK for audio processing with MLX on Apple Silicon

    774+5Star change over the last 7 days
  • BiBi-Keyboard@BryceWG

    说点啥(BiBi Keyboard):一个基于 Kotlin 的 Android 平台的 LLM 与 ASR 语音输入法键盘应用 An LLM ASR voice input method keyboard application for the Android platform based on Kotlin

    774+15Star change over the last 7 days
  • PaddlePaddle-DeepSpeech@yeyupiaoling

    基于PaddlePaddle实现的语音识别,中文语音识别。项目完善,识别效果好。支持Windows,Linux下训练和预测,支持Nvidia Jetson开发板预测。

    763+1Star change over the last 7 days
  • chaplin@amanvirparhar

    A real-time silent speech recognition tool.

    754+1Star change over the last 7 days
  • Running speech to text model (whisper.cpp) in Unity3d on your local machine.

    751+1Star change over the last 7 days
  • Transcribro@soupslurpr

    Private and on-device speech recognition keyboard and service for Android.

    745+2Star change over the last 7 days
← Back to topics