Skip to main content
buildradar
Sign in
Topic · voice-recognition

voice-recognition

Tracked open-source repos tagged voice-recognition, sorted by stars.

Repos
23
Total stars
85,588
Avg. stars
3,721
Share
0.00%

Topics that frequently appear alongside voice-recognition on the same repo.

Recent risers

Repos created in the last 90 days, tagged voice-recognition.

No new repos tagged with this topic in the last 90 days.

  • vosk-api@alphacep

    Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node

    15,101+22Star change over the last 7 days
  • PaddleSpeech@PaddlePaddle

    Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.

    12,676+6Star change over the last 7 days
  • speechbrain@speechbrain

    A PyTorch-based Speech Toolkit

    11,802+17Star change over the last 7 days
  • moonshine@moonshine-ai

    Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces

    10,997+53Star change over the last 7 days
  • silero-vad@snakers4

    Silero VAD: pre-trained enterprise-grade Voice Activity Detector

    10,112+48Star change over the last 7 days
  • WhisperLive@collabora

    A nearly-live implementation of OpenAI's Whisper.

    4,246+5Star change over the last 7 days
  • cnchar@theajack

    🇨🇳 功能全面的汉字工具库 (拼音 笔画 偏旁 成语 语音 可视化等) (Chinese character util)

    3,087+4Star change over the last 7 days
  • STT@coqui-ai

    🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.

    2,606+0Star change over the last 7 days
  • ten-vad@TEN-framework

    Voice Activity Detector (VAD) : low-latency, high-performance and lightweight

    2,256+8Star change over the last 7 days
  • voice_datasets@jim-schwoebel

    🔊 A comprehensive list of open-source datasets for voice and sound computing (95+ datasets).

    2,223+1Star change over the last 7 days
  • 💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies

    1,399+0Star change over the last 7 days
  • This project uses a variety of advanced voiceprint recognition models such as EcapaTdnn, ResNetSE, ERes2Net, CAM++, etc. It is not excluded that more models will be supported in the future. At the same time, this project also supports MelSpectrogram, Spectrogram data preprocessing methods

    1,313+1Star change over the last 7 days
  • Python AI assistant 🧠

    1,014+0Star change over the last 7 days
  • EDDiscovery@EDDiscovery

    Captains log and 3d star map for Elite Dangerous

    898+0Star change over the last 7 days
  • rhino@Picovoice

    On-device Speech-to-Intent engine powered by deep learning

    711+1Star change over the last 7 days
  • voicetypr@moinulmoin

    Voicetypr - AI powered offline voice to text dictation tool for busy founders, vibe coders, AI power users on macos, windows. Alternative to wispr flow and superwhisper.

    706+10Star change over the last 7 days
  • speech to text benchmark framework

    697+0Star change over the last 7 days
  • Speech Recognition for React Native Expo projects

    676+6Star change over the last 7 days
  • cheetah@Picovoice

    On-device streaming speech-to-text engine powered by deep learning

    671+0Star change over the last 7 days
  • sonus@evancohen

    :speech_balloon: /so.nus/ STT (speech to text) for Node with offline hotword detection

    638+0Star change over the last 7 days
  • LiveTranslate@TheDeathDragon

    Real-time audio translation, captures system audio + mic, runs ASR (Whisper/SenseVoice), translates via LLM API with streaming display. Perfect for VTubers, livestreamers, and watching foreign content. Windows 实时音频翻译,ASR 语音识别后 LLM 流式翻译显示,适合 VTuber、主播和外语视频观看。

    615+18Star change over the last 7 days
  • speech-to-text@reriiasu

    Real-time transcription using faster-whisper

    614+0Star change over the last 7 days
  • 🗣 An overlay that gets your user’s voice permission and input as text in a customizable UI

    557+0Star change over the last 7 days
← Back to topics