Skip to main content
buildradar
Sign in
Topic · speaker-diarization

speaker-diarization

Tracked open-source repos tagged speaker-diarization, sorted by stars.

Repos
20
Total stars
98,208
Avg. stars
4,910
Share
0.01%

Topics that frequently appear alongside speaker-diarization on the same repo.

Recent risers

Repos created in the last 90 days, tagged speaker-diarization.

No new repos tagged with this topic in the last 90 days.

  • FunASR@modelscope

    Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.

    20,136+64Star change over the last 7 days
  • speechbrain@speechbrain

    A PyTorch-based Speech Toolkit

    11,802+10Star change over the last 7 days
  • WhisperLiveKit@QuentinFuxa

    Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.

    10,990+16Star change over the last 7 days
  • pyannote-audio@pyannote

    Neural building blocks for speaker diarization: speech activity detection, speaker change detection, overlapped speech detection, speaker embedding

    10,500+16Star change over the last 7 days
  • espnet@espnet

    End-to-End Speech Processing Toolkit

    9,949+3Star change over the last 7 days
  • argmax-oss-swift@argmaxinc

    On-device Speech AI for Apple Silicon

    6,353+8Star change over the last 7 days
  • whisper-diarization@MahmoudAshraf97

    Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper

    5,634+1Star change over the last 7 days
  • Whisper & Faster-Whisper standalone executables for those who don't want to bother with Python.

    3,171+2Star change over the last 7 days
  • 3D-Speaker@modelscope

    A Repository for Single- and Multi-modal Speaker Verification, Speaker Recognition and Speaker Diarization

    3,128+3Star change over the last 7 days
  • whisper-timestamped@linto-ai

    Multilingual Automatic Speech Recognition with word-level timestamps and confidence

    2,842+1Star change over the last 7 days
  • FluidAudio@FluidInference

    Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.

    2,723+13Star change over the last 7 days
  • diart@juanmc2005

    A python package to build AI-powered real-time audio applications

    2,024+1Star change over the last 7 days
  • A curated list of awesome Speaker Diarization papers, libraries, datasets, and other resources.

    1,894+1Star change over the last 7 days
  • Fun-ASR@QwenAudio

    Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR, vLLM, streaming, and llama.cpp runtimes.

    1,512+10Star change over the last 7 days
  • wespeaker@wenet-e2e

    Research and Production Oriented Speaker Verification, Recognition and Diarization Toolkit

    1,402+8Star change over the last 7 days
  • speech-swift@soniqo

    AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML

    1,161+4Star change over the last 7 days
  • transcriptionstream@transcriptionstream

    turnkey self-hosted offline transcription and diarization service with llm summary

    948+2Star change over the last 7 days
  • MioSub@corvo007

    一站式全自动字幕生成软件,下载、转录、翻译、压制全流程覆盖,无需人工介入 / One-stop automated subtitle generator. Handles downloading, transcription, translation, and hardcoding—zero human intervention required.

    810+3Star change over the last 7 days
  • pyannote-whisper@yinruiqing
    677+0Star change over the last 7 days
  • SpectralCluster@wq2012

    Python re-implementation of the (constrained) spectral clustering algorithms used in Google's speaker diarization papers.

    555+0Star change over the last 7 days
← Back to topics