Skip to main content
buildradar
Sign in
Topic · vad

vad

Tracked open-source repos tagged vad, sorted by stars.

Repos
13
Total stars
32,319
Avg. stars
2,486
Share
0.00%

Topics that frequently appear alongside vad on the same repo.

Recent risers

Repos created in the last 90 days, tagged vad.

No new repos tagged with this topic in the last 90 days.

  • silero-vad@snakers4

    Silero VAD: pre-trained enterprise-grade Voice Activity Detector

    10,112+29Star change over the last 7 days
  • ffsubsync@smacke

    Automagically synchronize subtitles with video.

    7,864+7Star change over the last 7 days
  • faster-whisper-GUI@CheshireCC

    faster_whisper GUI with PySide6

    2,995+4Star change over the last 7 days
  • FluidAudio@FluidInference

    Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.

    2,723+13Star change over the last 7 days
  • ten-vad@TEN-framework

    Voice Activity Detector (VAD) : low-latency, high-performance and lightweight

    2,256+7Star change over the last 7 days
  • sherpa-ncnn@k2-fsa

    Real-time speech recognition and voice activity detection (VAD) using next-gen Kaldi with ncnn without Internet connection. Support iOS, Android, Linux, macOS, Windows, Raspberry Pi, VisionFive2, LicheePi4A etc.

    1,777-1Star change over the last 7 days
  • whisper.net@sandrohanea

    Whisper.net. Speech to text made simple using Whisper Models

    941+1Star change over the last 7 days
  • auditok@amsehili

    An voice activity detection and audio segmentation tool

    860+1Star change over the last 7 days
  • FireRedASR2S@FireRedTeam

    A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/accents), English, code-switching, and both speech and singing ASR. FireRedVAD supports speech/singing/music in 100+ langs. FireRedLID supports 100+ langs and 20+ zh dialects. FireRedPunc supports zh and en.

    670+13Star change over the last 7 days
  • WhisperS2T@shashikg

    An Optimized Speech-to-Text Pipeline for the Whisper Model Supporting Multiple Inference Engine

    578+1Star change over the last 7 days
  • ICASSP-2023-24-Papers@DmitryRyumin

    ICASSP 2023-2024 Papers: A complete collection of influential and exciting research papers from the ICASSP 2023-24 conferences. Explore the latest advancements in acoustics, speech and signal processing. Code included. Star the repository to support the advancement of audio and signal processing!

    526+1Star change over the last 7 days
  • FireRedVAD@FireRedTeam

    A SOTA Industrial-Grade Voice Activity Detection & Audio Event Detection, supporting 100+ languages, outperforming Silero-VAD, TEN-VAD, FunASR-VAD and WebRTC-VAD

    521+6Star change over the last 7 days
  • android-vad@gkonovalov

    Android Voice Activity Detection (VAD) library. Supports WebRTC VAD GMM, Silero VAD DNN, Yamnet VAD DNN models.

    506+1Star change over the last 7 days
← Back to topics