Skip to main content
buildradar
Sign in
Topic · speech-recognition

speech-recognition

Tracked open-source repos tagged speech-recognition, sorted by stars.

156 repos
  • audio-ai-hub@BinWang28

    The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.

    953+3Star change over the last 7 days
  • transcriptionstream@transcriptionstream

    turnkey self-hosted offline transcription and diarization service with llm summary

    948+2Star change over the last 7 days
  • whisper.net@sandrohanea

    Whisper.net. Speech to text made simple using Whisper Models

    941+1Star change over the last 7 days
  • espresso@freewym

    Espresso: A Fast End-to-End Neural Speech Recognition Toolkit

    939+0Star change over the last 7 days
  • A python based desktop voice assistant capable of executing system-level commands, integrating speech recognition and text-to-speech, and handling asynchronous user interactions.

    919+13Star change over the last 7 days
  • TheWhisper@TheStageAI

    Optimized Whisper models for streaming and on-device use

    897+0Star change over the last 7 days
  • A talking LLM that runs on your own computer without needing the internet.

    891+2Star change over the last 7 days
  • speechpy@astorfi

    :speech_balloon: SpeechPy - A Library for Speech Processing and Recognition: http://speechpy.readthedocs.io/en/latest/

    883+0Star change over the last 7 days
  • PPASR@yeyupiaoling

    基于PaddlePaddle实现端到端中文语音识别,从入门到实战,超简单的入门案例,超实用的企业项目。支持当前最流行的DeepSpeech2、Conformer、Squeezeformer模型

    870-1Star change over the last 7 days
  • react-speech-recognition@JamesBrill

    💬Speech recognition for your React app

    842+1Star change over the last 7 days
  • CTCDecoder@githubharald

    Connectionist Temporal Classification (CTC) decoding algorithms: best path, beam search, lexicon search, prefix search, and token passing. Implemented in Python.

    836-1Star change over the last 7 days
  • whisper-playground@saharmor

    Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/

    835+0Star change over the last 7 days
  • kur@deepgram

    Descriptive Deep Learning

    823-1Star change over the last 7 days
  • vocalinux@VocaHQ

    Free, open-source, 100% offline voice dictation for Linux. Speak and type anywhere via whisper.cpp, Whisper & VOSK engines, GPU-accelerated, works on X11 + Wayland!

    806+17Star change over the last 7 days
  • lobe-tts@lobehub

    🎤 Lobe TTS - A high-quality & reliable TTS/STT library for Server and Browser

    805+1Star change over the last 7 days
  • Speech to Text to Speech. Song now playing. Sends text as OSC messages to VRChat to display on avatar. (STTTS) (Speech to TTS) (VRC STT System) (VTuber TTS)

    803+3Star change over the last 7 days
  • whisper.rn@mybigday

    React Native binding of whisper.cpp.

    800-1Star change over the last 7 days
  • june@mezbaul-h

    Local voice chatbot for engaging conversations, powered by Ollama, Hugging Face Transformers, and Coqui TTS Toolkit

    788+0Star change over the last 7 days
  • whisper_mic@mallorbc

    Project that allows one to use a microphone with OpenAI whisper.

    787+0Star change over the last 7 days
  • GigaAM@salute-developers

    Foundational Model for Speech Recognition Tasks

    786+14Star change over the last 7 days
  • SwiftWhisper@exPHAT

    🎤 The easiest way to transcribe audio in Swift

    785-1Star change over the last 7 days
  • NotelyVoice@Notely-Voice

    A 100% private AI voice transcription app that converts speech to text in 100+ languages. Built with Compose Multiplatform for Android & iOS using Whisper AI - no cloud uploads, all processing happens on-device for complete privacy.

    777+1Star change over the last 7 days
  • cn2an@Ailln

    📦 快速转化「中文数字」和「阿拉伯数字」~ (最新特性:分数,日期、温度等转化)

    766+0Star change over the last 7 days
  • dla@markovka17

    Deep learning for audio processing

    763+1Star change over the last 7 days
  • PaddlePaddle-DeepSpeech@yeyupiaoling

    基于PaddlePaddle实现的语音识别,中文语音识别。项目完善,识别效果好。支持Windows,Linux下训练和预测,支持Nvidia Jetson开发板预测。

    762+0Star change over the last 7 days
  • chaplin@amanvirparhar

    A real-time silent speech recognition tool.

    754+2Star change over the last 7 days
  • Running speech to text model (whisper.cpp) in Unity3d on your local machine.

    751+1Star change over the last 7 days
  • allosaurus@xinjli

    Allosaurus is a pretrained universal phone recognizer for more than 2000 languages

    746+3Star change over the last 7 days
  • Transcribro@soupslurpr

    Private and on-device speech recognition keyboard and service for Android.

    745+5Star change over the last 7 days
  • MASR@yeyupiaoling

    Pytorch实现的流式与非流式的自动语音识别框架,同时兼容在线和离线识别,目前支持Conformer、Squeezeformer、DeepSpeech2模型,支持多种数据增强方法。

    727+0Star change over the last 7 days
← Back to topics