speech-recognition
Tracked open-source repos tagged speech-recognition, sorted by stars.
- #31
JAX implementation of OpenAI's Whisper model for up to 70x speed-up on TPU.
★ 4,679-1Star change over the last 7 days - #32
A small speech recognizer
★ 4,337+2Star change over the last 7 days - #33
Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.
★ 4,114+2Star change over the last 7 days - #34
OpenAI Whisper ASR Webservice API
★ 3,328+2Star change over the last 7 days - #35
Whisper & Faster-Whisper standalone executables for those who don't want to bother with Python.
★ 3,171+2Star change over the last 7 days - #36
Open source, local, and self-hosted Amazon Echo/Google Home competitive Voice Assistant alternative
★ 3,101+1Star change over the last 7 days - #37
Transcribe any audio to text, translate and edit subtitles 100% locally with a web UI. Powered by whisper models!
★ 3,068+2Star change over the last 7 days - #38★ 2,864+0Star change over the last 7 days
- #39
Multilingual Automatic Speech Recognition with word-level timestamps and confidence
★ 2,842+1Star change over the last 7 days - #40
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
★ 2,606+1Star change over the last 7 days - #41
A realtime voice runtime that keeps Agents talking, working, and present. Real-time Voice Runtime for AI Agents
★ 2,362+68Star change over the last 7 days - #42
🎙Speech recognition using the tensorflow deep learning framework, sequence-to-sequence neural networks
★ 2,173+0Star change over the last 7 days - #43★ 2,048+7Star change over the last 7 days
- #44
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering outstanding singing lyrics recognition capability.
★ 1,975+2Star change over the last 7 days - #45★ 1,968+0Star change over the last 7 days
- #46★ 1,935+1Star change over the last 7 days
- #47
A curated list of awesome Speaker Diarization papers, libraries, datasets, and other resources.
★ 1,894+1Star change over the last 7 days - #48
The Self-Coding System for Your App — Alan AI SDK for iOS
★ 1,878-3Star change over the last 7 days - #49
The Self-Coding System for Your App — Alan AI SDK for Android
★ 1,807-1Star change over the last 7 days - #50
Cross-Platform, GPU Accelerated Whisper 🏎️
★ 1,794+0Star change over the last 7 days - #51
Real-time speech recognition and voice activity detection (VAD) using next-gen Kaldi with ncnn without Internet connection. Support iOS, Android, Linux, macOS, Windows, Raspberry Pi, VisionFive2, LicheePi4A etc.
★ 1,777-1Star change over the last 7 days - #52
The Self-Coding System for Your App — Alan AI SDK for Flutter
★ 1,756-1Star change over the last 7 days - #53
The Self-Coding System for Your App — Alan AI SDK for Ionic
★ 1,649-1Star change over the last 7 days - #54
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.
★ 1,622+8Star change over the last 7 days - #55
OBS plugin for local speech recognition and captioning using AI
★ 1,597+8Star change over the last 7 days - #56
Amica is an open source interface for interactive communication with 3D characters with voice synthesis and speech recognition.
★ 1,591-2Star change over the last 7 days - #57
Custom nodes that extend the capabilities of Comfyui
★ 1,525+1Star change over the last 7 days - #58★ 1,521+3Star change over the last 7 days
- #59
Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR, vLLM, streaming, and llama.cpp runtimes.
★ 1,512+10Star change over the last 7 days - #60★ 1,449+0Star change over the last 7 days