voice-recognition
Tracked open-source repos tagged voice-recognition, sorted by stars.
Related topics
Topics that frequently appear alongside voice-recognition on the same repo.
Recent risers
Repos created in the last 90 days, tagged voice-recognition.
No new repos tagged with this topic in the last 90 days.
- #1
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
★ 15,101+22Star change over the last 7 days - #2
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.
★ 12,676+6Star change over the last 7 days - #3
A PyTorch-based Speech Toolkit
★ 11,802+17Star change over the last 7 days - #4
Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces
★ 10,997+53Star change over the last 7 days - #5
Silero VAD: pre-trained enterprise-grade Voice Activity Detector
★ 10,112+48Star change over the last 7 days - #6
A nearly-live implementation of OpenAI's Whisper.
★ 4,246+5Star change over the last 7 days - #7★ 3,087+4Star change over the last 7 days
- #8
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
★ 2,606+0Star change over the last 7 days - #9★ 2,256+8Star change over the last 7 days
- #10
🔊 A comprehensive list of open-source datasets for voice and sound computing (95+ datasets).
★ 2,223+1Star change over the last 7 days - #11
💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
★ 1,399+0Star change over the last 7 days - #12
This project uses a variety of advanced voiceprint recognition models such as EcapaTdnn, ResNetSE, ERes2Net, CAM++, etc. It is not excluded that more models will be supported in the future. At the same time, this project also supports MelSpectrogram, Spectrogram data preprocessing methods
★ 1,313+1Star change over the last 7 days - #13
Python AI assistant 🧠
★ 1,014+0Star change over the last 7 days - #14
Captains log and 3d star map for Elite Dangerous
★ 898+0Star change over the last 7 days - #15★ 711+1Star change over the last 7 days
- #16
Voicetypr - AI powered offline voice to text dictation tool for busy founders, vibe coders, AI power users on macos, windows. Alternative to wispr flow and superwhisper.
★ 706+10Star change over the last 7 days - #17
speech to text benchmark framework
★ 697+0Star change over the last 7 days - #18
Speech Recognition for React Native Expo projects
★ 676+6Star change over the last 7 days - #19★ 671+0Star change over the last 7 days
- #20
:speech_balloon: /so.nus/ STT (speech to text) for Node with offline hotword detection
★ 638+0Star change over the last 7 days - #21
Real-time audio translation, captures system audio + mic, runs ASR (Whisper/SenseVoice), translates via LLM API with streaming display. Perfect for VTubers, livestreamers, and watching foreign content. Windows 实时音频翻译,ASR 语音识别后 LLM 流式翻译显示,适合 VTuber、主播和外语视频观看。
★ 615+18Star change over the last 7 days - #22
Real-time transcription using faster-whisper
★ 614+0Star change over the last 7 days - #23
🗣 An overlay that gets your user’s voice permission and input as text in a customizable UI
★ 557+0Star change over the last 7 days