speaker-diarization
Tracked open-source repos tagged speaker-diarization, sorted by stars.
Related topics
Topics that frequently appear alongside speaker-diarization on the same repo.
Recent risers
Repos created in the last 90 days, tagged speaker-diarization.
No new repos tagged with this topic in the last 90 days.
- #1
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
★ 20,136+64Star change over the last 7 days - #2
A PyTorch-based Speech Toolkit
★ 11,802+10Star change over the last 7 days - #3
Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
★ 10,990+16Star change over the last 7 days - #4
Neural building blocks for speaker diarization: speech activity detection, speaker change detection, overlapped speech detection, speaker embedding
★ 10,500+16Star change over the last 7 days - #5★ 9,949+3Star change over the last 7 days
- #6
On-device Speech AI for Apple Silicon
★ 6,353+8Star change over the last 7 days - #7
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
★ 5,634+1Star change over the last 7 days - #8
Whisper & Faster-Whisper standalone executables for those who don't want to bother with Python.
★ 3,171+2Star change over the last 7 days - #9
A Repository for Single- and Multi-modal Speaker Verification, Speaker Recognition and Speaker Diarization
★ 3,128+3Star change over the last 7 days - #10
Multilingual Automatic Speech Recognition with word-level timestamps and confidence
★ 2,842+1Star change over the last 7 days - #11
Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.
★ 2,723+13Star change over the last 7 days - #12★ 2,024+1Star change over the last 7 days
- #13
A curated list of awesome Speaker Diarization papers, libraries, datasets, and other resources.
★ 1,894+1Star change over the last 7 days - #14
Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR, vLLM, streaming, and llama.cpp runtimes.
★ 1,512+10Star change over the last 7 days - #15
Research and Production Oriented Speaker Verification, Recognition and Diarization Toolkit
★ 1,402+8Star change over the last 7 days - #16
AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
★ 1,161+4Star change over the last 7 days - #17
turnkey self-hosted offline transcription and diarization service with llm summary
★ 948+2Star change over the last 7 days - #18
一站式全自动字幕生成软件,下载、转录、翻译、压制全流程覆盖,无需人工介入 / One-stop automated subtitle generator. Handles downloading, transcription, translation, and hardcoding—zero human intervention required.
★ 810+3Star change over the last 7 days - #19★ 677+0Star change over the last 7 days
- #20
Python re-implementation of the (constrained) spectral clustering algorithms used in Google's speaker diarization papers.
★ 555+0Star change over the last 7 days