speech-to-text
Tracked open-source repos tagged speech-to-text, sorted by stars.
- #91
面向抖音直播电商的 Windows 本地 AI Agent Studio,贯通主播发现、直播洞察、直播复盘与短视频内容编导的统一智能工作流。
★ 1,018+424Star change over the last 7 days - #92
:zap: TensorFlowASR: Almost State-of-the-art Automatic Speech Recognition in Tensorflow 2. Supported languages that can use characters or subwords
★ 1,010+0Star change over the last 7 days - #93★ 1,006-1Star change over the last 7 days
- #94
Whisper-Flow is a framework designed to enable real-time transcription of audio content using OpenAI’s Whisper model. Rather than processing entire files after upload (“batch mode”), Whisper-Flow accepts a continuous stream of audio chunks and produces incremental transcripts immediately.
★ 964+25Star change over the last 7 days - #95
Near-Realtime audio transcription using self-hosted Whisper and WebSocket in Python/JS
★ 960+0Star change over the last 7 days - #96
Botium Speech Processing
★ 943+0Star change over the last 7 days - #97
Whisper.net. Speech to text made simple using Whisper Models
★ 942+1Star change over the last 7 days - #98
Evaluate your speech-to-text system with similarity measures such as word error rate (WER)
★ 928+2Star change over the last 7 days - #99
.NET MAUI Samples
★ 915+0Star change over the last 7 days - #100★ 911+2Star change over the last 7 days
- #101
Optimized Whisper models for streaming and on-device use
★ 897-1Star change over the last 7 days - #102★ 896+4Star change over the last 7 days
- #103
基于PaddlePaddle实现端到端中文语音识别,从入门到实战,超简单的入门案例,超实用的企业项目。支持当前最流行的DeepSpeech2、Conformer、Squeezeformer模型
★ 870-1Star change over the last 7 days - #104
💬Speech recognition for your React app
★ 842+0Star change over the last 7 days - #105
Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/
★ 835+0Star change over the last 7 days - #106★ 823-1Star change over the last 7 days
- #107
一站式全自动字幕生成软件,下载、转录、翻译、压制全流程覆盖,无需人工介入 / One-stop automated subtitle generator. Handles downloading, transcription, translation, and hardcoding—zero human intervention required.
★ 811+3Star change over the last 7 days - #108
Free, open-source, 100% offline voice dictation for Linux. Speak and type anywhere via whisper.cpp, Whisper & VOSK engines, GPU-accelerated, works on X11 + Wayland!
★ 809+17Star change over the last 7 days - #109★ 805-1Star change over the last 7 days
- #110
Speech to Text to Speech. Song now playing. Sends text as OSC messages to VRChat to display on avatar. (STTTS) (Speech to TTS) (VRC STT System) (VTuber TTS)
★ 804+5Star change over the last 7 days - #111
Local voice chatbot for engaging conversations, powered by Ollama, Hugging Face Transformers, and Coqui TTS Toolkit
★ 788+1Star change over the last 7 days - #112
Project that allows one to use a microphone with OpenAI whisper.
★ 787+0Star change over the last 7 days - #113
🎤 The easiest way to transcribe audio in Swift
★ 785+0Star change over the last 7 days - #114
A 100% private AI voice transcription app that converts speech to text in 100+ languages. Built with Compose Multiplatform for Android & iOS using Whisper AI - no cloud uploads, all processing happens on-device for complete privacy.
★ 779+4Star change over the last 7 days - #115
A modular Swift SDK for audio processing with MLX on Apple Silicon
★ 774+5Star change over the last 7 days - #116
说点啥(BiBi Keyboard):一个基于 Kotlin 的 Android 平台的 LLM 与 ASR 语音输入法键盘应用 An LLM ASR voice input method keyboard application for the Android platform based on Kotlin
★ 774+15Star change over the last 7 days - #117
基于PaddlePaddle实现的语音识别,中文语音识别。项目完善,识别效果好。支持Windows,Linux下训练和预测,支持Nvidia Jetson开发板预测。
★ 763+1Star change over the last 7 days - #118★ 754+1Star change over the last 7 days
- #119
Running speech to text model (whisper.cpp) in Unity3d on your local machine.
★ 751+1Star change over the last 7 days - #120
Private and on-device speech recognition keyboard and service for Android.
★ 745+2Star change over the last 7 days