Skip to main content
buildradar
Sign in
Topic · speech-processing

speech-processing

Tracked open-source repos tagged speech-processing, sorted by stars.

Repos
20
Total stars
71,137
Avg. stars
3,557
Share
0.00%

Topics that frequently appear alongside speech-processing on the same repo.

Recent risers

Repos created in the last 90 days, tagged speech-processing.

No new repos tagged with this topic in the last 90 days.

  • speechbrain@speechbrain

    A PyTorch-based Speech Toolkit

    11,792+22Star change over the last 7 days
  • pyannote-audio@pyannote

    Neural building blocks for speaker diarization: speech activity detection, speaker change detection, overlapped speech detection, speaker embedding

    10,484+38Star change over the last 7 days
  • silero-vad@snakers4

    Silero VAD: pre-trained enterprise-grade Voice Activity Detector

    10,083+52Star change over the last 7 days
  • maths-cs-ai-compendium@HenryNdubuaku

    Become a cracked AI/ML researcher/engineer with this unconventional textbook covering maths, computing, and ML with intuition.

    7,379+33Star change over the last 7 days
  • Reading list for research topics in multimodal machine learning

    6,925+0Star change over the last 7 days
  • torchscale@microsoft

    Foundation Architecture for (M)LLMs

    3,138+1Star change over the last 7 days
  • whisper-timestamped@linto-ai

    Multilingual Automatic Speech Recognition with word-level timestamps and confidence

    2,841+2Star change over the last 7 days
  • resemble-enhance@resemble-ai

    AI powered speech denoising and enhancement

    2,401+6Star change over the last 7 days
  • ten-vad@TEN-framework

    Voice Activity Detector (VAD) : low-latency, high-performance and lightweight

    2,250+4Star change over the last 7 days
  • IMS-Toucan@DigitalPhonetics

    Controllable and fast Text-to-Speech for over 7000 languages!

    2,208+0Star change over the last 7 days
  • A curated list of awesome Speaker Diarization papers, libraries, datasets, and other resources.

    1,893+2Star change over the last 7 days
  • 💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies

    1,399+0Star change over the last 7 days
  • voicefixer@haoheliu

    General Speech Restoration

    1,376+3Star change over the last 7 days
  • CrisperWhisper@nyrahealth

    Controllable Transcription. Verbatim ( every, filler, pause, stutter, vocal sound) , or intended ( what the speaker meant to say, optimized for readability) with word-level timestamps.

    1,362+21Star change over the last 7 days
  • StreamSpeech@ictnlp

    StreamSpeech is an “All in One” seamless model for offline and simultaneous speech recognition, speech translation and speech synthesis.

    1,288+2Star change over the last 7 days
  • audino@midas-research

    Open source audio annotation tool for humans

    1,146+0Star change over the last 7 days
  • SLAM-LLM@X-LANCE

    A Framework for Speech, Language, Audio, Music Processing with Large Language Model

    1,056+0Star change over the last 7 days
  • You can find the speech algorithms you want here

    878+1Star change over the last 7 days
  • MultiBench@pliang279

    [NeurIPS 2021] Multiscale Benchmarks for Multimodal Representation Learning

    636+0Star change over the last 7 days
  • 语音方向实验室/公司/资源/实习等,欢迎推荐或自荐

    609+0Star change over the last 7 days
← Back to topics