Skip to main content
buildradar
Sign in
Topic · transcription

transcription

Tracked open-source repos tagged transcription, sorted by stars.

Repos
53
Total stars
175,951
Avg. stars
3,320
Share
0.01%

Topics that frequently appear alongside transcription on the same repo.

Recent risers

Repos created in the last 90 days, tagged transcription.

  • claude-real-video@HUANGCHIHHUNGLeo

    Let Claude (or any LLM) actually watch a video — scene-aware, deduplicated frames + transcript, from a URL or local file. Runs locally, MIT.

    2,109
  • rescript@wassgha

    🎬 Open source, transcript-based video/audio editor that lives in the browser.

    866
  • meetily@Zackriya-Solutions

    Privacy first, AI meeting assistant with 4x faster Parakeet/Whisper live transcription, speaker diarization, and Ollama summarization built on Rust. 100% local processing. no cloud required. Meetily (Meetly Ai - https://meetily.ai) is the #1 Self-hosted, Open-source Ai meeting note taker for macOS & Windows. Understand How to write meeting minutes

    30,252+195Star change over the last 7 days
  • FunASR@modelscope

    Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.

    20,136+64Star change over the last 7 days
  • VoiceStudio@debpalash

    VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.

    14,835+2,880Star change over the last 7 days
  • omi@BasedHardware

    AI that sees your screen, listens to your conversations and tells you what to do

    13,377+46Star change over the last 7 days
  • voice-pro@abus-aikorea

    Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

    12,730+54Star change over the last 7 days
  • SenseVoice@QwenAudio

    Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

    9,209+37Star change over the last 7 days
  • FunClip@modelscope

    FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.

    6,210+13Star change over the last 7 days
  • basic-pitch@spotify

    A lightweight yet powerful audio-to-MIDI converter with pitch bend detection

    5,522+21Star change over the last 7 days
  • auto-subs@tmoroney

    On-device subtitle generation that connects directly to DaVinci Resolve, Premiere, and After Effects.

    4,122+20Star change over the last 7 days
  • speaches@speaches-ai
    3,638+8Star change over the last 7 days
  • whishper@pluja

    Transcribe any audio to text, translate and edit subtitles 100% locally with a web UI. Powered by whisper models!

    3,068+2Star change over the last 7 days
  • Scriberr@rishikanthc

    Self-hosted AI audio transcription

    3,016+14Star change over the last 7 days
  • Hex@kitlangton

    Legacy Swift Hex app. Try the Rust rewrite at hex.kitlangton.com; new source at github.com/anomalyco/hex.

    2,893+2Star change over the last 7 days
  • vexa@Vexa-ai

    Open-source meeting transcription API for Google Meet, Microsoft Teams & Zoom. Auto-join bots, real-time WebSocket transcripts, MCP server for AI agents. Self-host or use hosted SaaS.

    2,741+15Star change over the last 7 days
  • awesome-whisper@sindresorhus

    🔊 Awesome list for Whisper — an open-source AI-powered speech recognition system developed by OpenAI

    2,375+1Star change over the last 7 days
  • kalosm@floneum

    Instant, controllable, local pre-trained AI models in Rust

    2,226+3Star change over the last 7 days
  • noScribe@kaixxx

    Cutting edge AI technology for automated audio transcription. A nice GUI for OpenAIs Whisper and pyannote (speaker identification)

    2,145+8Star change over the last 7 days
  • claude-real-video@HUANGCHIHHUNGLeo

    Let Claude (or any LLM) actually watch a video — scene-aware, deduplicated frames + transcript, from a URL or local file. Runs locally, MIT.

    2,109+20Star change over the last 7 days
  • diart@juanmc2005

    A python package to build AI-powered real-time audio applications

    2,024+1Star change over the last 7 days
  • audapolis@bugbakery

    an editor for spoken-word audio with automatic transcription

    1,892+2Star change over the last 7 days
  • typewhisper-mac@TypeWhisper

    Local speech-to-text for macOS on-device AI, fully private, optional cloud

    1,759+12Star change over the last 7 days
  • OBS plugin for local speech recognition and captioning using AI

    1,597+8Star change over the last 7 days
  • amical@amicalhq

    🎙️ AI Dictation App - Open Source and Local-first ⚡ Type 3x faster, no keyboard needed. 🆓 Powered by open source models, works offline, fast and accurate.

    1,520+8Star change over the last 7 days
  • pianotrans@azuwis

    Simple GUI for ByteDance's Piano Transcription with Pedals

    1,514+3Star change over the last 7 days
  • Fun-ASR@QwenAudio

    Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR, vLLM, streaming, and llama.cpp runtimes.

    1,512+10Star change over the last 7 days
  • minutes@silverstein

    Open-source, local-first Granola/Otter alternative that Claude Code, Codex, Cursor, and any MCP client can query. Meetings, calls, and voice memos transcribed on-device into markdown you own.

    1,462+6Star change over the last 7 days
  • Multi-backend whisper app. Blazing fast. Mac-arm optimized. Easy install. Input a local file or url and this service will transcribe it using Whisper AI. Completely private and Free 🤯🤯🤯

    1,394+0Star change over the last 7 days
  • CrisperWhisper@nyrahealth

    Controllable Transcription. Verbatim ( every, filler, pause, stutter, vocal sound) , or intended ( what the speaker meant to say, optimized for readability) with word-level timestamps.

    1,371+13Star change over the last 7 days
  • book@hardhackerlabs

    「硬地骇客 - 两个月 $12000 ARR 实践之路」是由 硬地骇客 团队编著,本书是关于 Podwise 产品历程的忠实记录:内容包含 灵感 - 构建 - 发布 - 增长 - 复盘 五个章节。如果你觉得一个人读不够过瘾,欢迎加入「硬地骇客」官方知识星球与专家们一起讨论!Podwise 的故事才刚刚开始,我们也将在星球持续分享我们的认知,成功可能无法复制,但失败一定可以借鉴。现在就点击下方链接加入吧!

    1,370+0Star change over the last 7 days
  • 视频音频生成字幕,生成srt文件。无需申请第三方API,本地实现音频转文本。基于Transformer的视频字幕生成框架。A GUI tool for generating subtitle from videos and generating srt files.

    1,179+3Star change over the last 7 days
← Back to topics