automatic-speech-recognition
Tracked open-source repos tagged automatic-speech-recognition, sorted by stars.
Related topics
Topics that frequently appear alongside automatic-speech-recognition on the same repo.
Recent risers
Repos created in the last 90 days, tagged automatic-speech-recognition.
No new repos tagged with this topic in the last 90 days.
- #1
Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
★ 10,990+26Star change over the last 7 days - #2★ 5,229+2Star change over the last 7 days
- #3
OpenAI Whisper ASR Webservice API
★ 3,328+2Star change over the last 7 days - #4
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
★ 3,061+283Star change over the last 7 days - #5
Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.
★ 2,723+22Star change over the last 7 days - #6
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
★ 2,606+0Star change over the last 7 days - #7★ 2,256+8Star change over the last 7 days
- #8
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering outstanding singing lyrics recognition capability.
★ 1,975+3Star change over the last 7 days - #9
:zap: TensorFlowASR: Almost State-of-the-art Automatic Speech Recognition in Tensorflow 2. Supported languages that can use characters or subwords
★ 1,010+0Star change over the last 7 days - #10
Evaluate your speech-to-text system with similarity measures such as word error rate (WER)
★ 927+1Star change over the last 7 days - #11
Collection of resources on the applications of Large Language Models (LLMs) in Audio AI.
★ 738+1Star change over the last 7 days - #12
Offline Speech Recognition with OpenAI Whisper and TensorFlow Lite for Android
★ 688-1Star change over the last 7 days - #13★ 671+0Star change over the last 7 days
- #14
A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/accents), English, code-switching, and both speech and singing ASR. FireRedVAD supports speech/singing/music in 100+ langs. FireRedLID supports 100+ langs and 20+ zh dialects. FireRedPunc supports zh and en.
★ 669+10Star change over the last 7 days