speech-recognition
Tracked open-source repos tagged speech-recognition, sorted by stars.
- #91
The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
★ 953+3Star change over the last 7 days - #92
turnkey self-hosted offline transcription and diarization service with llm summary
★ 948+2Star change over the last 7 days - #93
Whisper.net. Speech to text made simple using Whisper Models
★ 941+1Star change over the last 7 days - #94★ 939+0Star change over the last 7 days
- #95
A python based desktop voice assistant capable of executing system-level commands, integrating speech recognition and text-to-speech, and handling asynchronous user interactions.
★ 919+13Star change over the last 7 days - #96
Optimized Whisper models for streaming and on-device use
★ 897+0Star change over the last 7 days - #97
A talking LLM that runs on your own computer without needing the internet.
★ 891+2Star change over the last 7 days - #98
:speech_balloon: SpeechPy - A Library for Speech Processing and Recognition: http://speechpy.readthedocs.io/en/latest/
★ 883+0Star change over the last 7 days - #99
基于PaddlePaddle实现端到端中文语音识别,从入门到实战,超简单的入门案例,超实用的企业项目。支持当前最流行的DeepSpeech2、Conformer、Squeezeformer模型
★ 870-1Star change over the last 7 days - #100
💬Speech recognition for your React app
★ 842+1Star change over the last 7 days - #101
Connectionist Temporal Classification (CTC) decoding algorithms: best path, beam search, lexicon search, prefix search, and token passing. Implemented in Python.
★ 836-1Star change over the last 7 days - #102
Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/
★ 835+0Star change over the last 7 days - #103★ 823-1Star change over the last 7 days
- #104
Free, open-source, 100% offline voice dictation for Linux. Speak and type anywhere via whisper.cpp, Whisper & VOSK engines, GPU-accelerated, works on X11 + Wayland!
★ 806+17Star change over the last 7 days - #105★ 805+1Star change over the last 7 days
- #106
Speech to Text to Speech. Song now playing. Sends text as OSC messages to VRChat to display on avatar. (STTTS) (Speech to TTS) (VRC STT System) (VTuber TTS)
★ 803+3Star change over the last 7 days - #107
React Native binding of whisper.cpp.
★ 800-1Star change over the last 7 days - #108
Local voice chatbot for engaging conversations, powered by Ollama, Hugging Face Transformers, and Coqui TTS Toolkit
★ 788+0Star change over the last 7 days - #109
Project that allows one to use a microphone with OpenAI whisper.
★ 787+0Star change over the last 7 days - #110★ 786+14Star change over the last 7 days
- #111
🎤 The easiest way to transcribe audio in Swift
★ 785-1Star change over the last 7 days - #112
A 100% private AI voice transcription app that converts speech to text in 100+ languages. Built with Compose Multiplatform for Android & iOS using Whisper AI - no cloud uploads, all processing happens on-device for complete privacy.
★ 777+1Star change over the last 7 days - #113★ 766+0Star change over the last 7 days
- #114★ 763+1Star change over the last 7 days
- #115
基于PaddlePaddle实现的语音识别,中文语音识别。项目完善,识别效果好。支持Windows,Linux下训练和预测,支持Nvidia Jetson开发板预测。
★ 762+0Star change over the last 7 days - #116★ 754+2Star change over the last 7 days
- #117
Running speech to text model (whisper.cpp) in Unity3d on your local machine.
★ 751+1Star change over the last 7 days - #118
Allosaurus is a pretrained universal phone recognizer for more than 2000 languages
★ 746+3Star change over the last 7 days - #119
Private and on-device speech recognition keyboard and service for Android.
★ 745+5Star change over the last 7 days - #120
Pytorch实现的流式与非流式的自动语音识别框架,同时兼容在线和离线识别,目前支持Conformer、Squeezeformer、DeepSpeech2模型,支持多种数据增强方法。
★ 727+0Star change over the last 7 days