Skip to main content
buildradar
Sign in
Topic · stt

stt

Tracked open-source repos tagged stt, sorted by stars.

Repos
32
Total stars
109,255
Avg. stars
3,414
Share
0.01%

Topics that frequently appear alongside stt on the same repo.

Recent risers

Repos created in the last 90 days, tagged stt.

  • floe-guard@Floe-Labs

    The spend meter and budget gate for AI voice agents. Meters STT + TTS + LLM + telephony per call, out of the box (Pipecat, LiveKit — Python & TypeScript). Hard-stops the next turn before it crosses your ceiling. Local, no account, no telemetry. Built by Floe — cost controls for Voice AI.

    434
  • khoj@khoj-ai

    Your AI second brain. Self-hostable. Get answers from the web or your docs. Build custom agents, schedule automations, do deep research. Turn any online or local LLM into your personal, autonomous AI (gpt, claude, gemini, llama, qwen, mistral). Get started - free.

    37,010+223Star change over the last 7 days
  • vosk-api@alphacep

    Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node

    15,101+17Star change over the last 7 days
  • moonshine@moonshine-ai

    Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces

    10,997+33Star change over the last 7 days
  • Vision-Agents@GetStream

    Open Vision Agents by Stream. Build voice and vision agents quickly with any model or video provider. Uses Stream's edge network for ultra-low latency.

    8,113+9Star change over the last 7 days
  • stt@jianchang512

    Voice Recognition to Text Tool / 一个离线运行的本地音视频转字幕工具,输出json、srt字幕、纯文字格式

    4,780+9Star change over the last 7 days
  • whishper@pluja

    Transcribe any audio to text, translate and edit subtitles 100% locally with a web UI. Powered by whisper models!

    3,068+2Star change over the last 7 days
  • STT@coqui-ai

    🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.

    2,606+1Star change over the last 7 days
  • 🎙Speech recognition using the tensorflow deep learning framework, sequence-to-sequence neural networks

    2,173+0Star change over the last 7 days
  • ElatoAI@akdeb

    Realtime Voice AI with 100+ Models on Arduino ESP32 with Secure Websockets and Edge Functions for AI Companions, and Devices

    1,938+5Star change over the last 7 days
  • Meet Ava, the WhatsApp Agent

    1,675+2Star change over the last 7 days
  • dsnote@mkiol

    Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.

    1,622+8Star change over the last 7 days
  • dicio-android@DicioTeam

    Dicio assistant app for Android

    1,463+3Star change over the last 7 days
  • Speech-AI-Forge@lenML

    🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.

    1,418+0Star change over the last 7 days
  • SoniTranslate@R3gm

    Synchronized Translation for Videos. Video dubbing

    1,413+2Star change over the last 7 days
  • 💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies

    1,399+0Star change over the last 7 days
  • 小智ESP32的Java企业级管理平台,提供设备监控、音色定制、角色切换和对话记录管理的前后端及服务端一体化解决方案

    1,337+1Star change over the last 7 days
  • gp.nvim@Robitx

    Gp.nvim (GPT prompt) Neovim AI plugin: ChatGPT sessions & Instructable text/code operations & Speech to text [OpenAI, Ollama, Anthropic, ..]

    1,320+0Star change over the last 7 days
  • my-translator@phuc-nt

    Real-time speech translation — macOS & Windows, free TTS, no server, your API keys only

    1,278+2Star change over the last 7 days
  • sokuji@kizuna-ai-lab

    Real-time two-way speech translation for bilingual meetings — auto-detects the spoken language and translates both directions, cloud or fully offline on-device. Desktop (Windows · macOS · Linux) + browser extension (Chrome · Edge) for Zoom, Meet, Teams & any app.

    1,186+40Star change over the last 7 days
  • nobodywho@nobodywho-ooo

    NobodyWho is an inference engine that lets you run LLMs locally and efficiently on any device.

    1,094+13Star change over the last 7 days
  • lobe-tts@lobehub

    🎤 Lobe TTS - A high-quality & reliable TTS/STT library for Server and Browser

    805+1Star change over the last 7 days
  • voxt@hehehai

    🎙️ An intelligent voice productivity assistant that turns speech into clean text, useful actions, and structured knowledge. It helps users capture ideas, communicate naturally, automate repetitive tasks, and stay productive across different apps and workflows.

    804+3Star change over the last 7 days
  • Speech to Text to Speech. Song now playing. Sends text as OSC messages to VRChat to display on avatar. (STTTS) (Speech to TTS) (VRC STT System) (VTuber TTS)

    802+3Star change over the last 7 days
  • mlx-audio-swift@Blaizzy

    A modular Swift SDK for audio processing with MLX on Apple Silicon

    772+6Star change over the last 7 days
  • Running speech to text model (whisper.cpp) in Unity3d on your local machine.

    751+1Star change over the last 7 days
  • mlx-omni-server@madroidmaq

    MLX Omni Server is a local inference server powered by Apple's MLX framework, specifically designed for Apple Silicon (M-series) chips. It implements OpenAI-compatible API endpoints, enabling seamless integration with existing OpenAI SDK clients while leveraging the power of local ML inference.

    741+1Star change over the last 7 days
  • cheetah@Picovoice

    On-device streaming speech-to-text engine powered by deep learning

    671+0Star change over the last 7 days
  • sonus@evancohen

    :speech_balloon: /so.nus/ STT (speech to text) for Node with offline hotword detection

    638+0Star change over the last 7 days
  • A React component to make correcting automated transcriptions of audio and video easier and faster. By BBC News Labs. - Work in progress

    620-1Star change over the last 7 days
  • CrispASR@CrispStrobe

    C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more

    610+21Star change over the last 7 days
← Back to topics