transcription
Tracked open-source repos tagged transcription, sorted by stars.
- #31
Muesli: agent-native local meeting transcription + dictation for macOS (Granola + WisprFlow alternative)
★ 1,105+35Star change over the last 7 days - #32
The open-source iOS app that's making quality voice transcription more accessible on mobile devices.
★ 1,102+6Star change over the last 7 days - #33
Whisper-Flow is a framework designed to enable real-time transcription of audio content using OpenAI’s Whisper model. Rather than processing entire files after upload (“batch mode”), Whisper-Flow accepts a continuous stream of audio chunks and produces incremental transcripts immediately.
★ 964+25Star change over the last 7 days - #34
turnkey self-hosted offline transcription and diarization service with llm summary
★ 948+1Star change over the last 7 days - #35
Optimized Whisper models for streaming and on-device use
★ 897-1Star change over the last 7 days - #36★ 868+0Star change over the last 7 days
- #37★ 868+13Star change over the last 7 days
- #38
一站式全自动字幕生成软件,下载、转录、翻译、压制全流程覆盖,无需人工介入 / One-stop automated subtitle generator. Handles downloading, transcription, translation, and hardcoding—zero human intervention required.
★ 811+3Star change over the last 7 days - #39
🎤 The easiest way to transcribe audio in Swift
★ 785+0Star change over the last 7 days - #40
Easily take an entire YouTube playlist and turn it into high quality transcripts using Whisper.
★ 689+2Star change over the last 7 days - #41
Offline Speech Recognition with OpenAI Whisper and TensorFlow Lite for Android
★ 688+0Star change over the last 7 days - #42★ 671+0Star change over the last 7 days
- #43
Command line interface for the built-in speech recognition and transcription capabilities in macOS.
★ 669+0Star change over the last 7 days - #44
把中文全渠道内容(抖音 / B站 / 小红书 / 公众号 / X / 播客)采集进个人知识库的 13 个 AI Skill:图文存图、视频转文字稿、字幕优先免 GPU,附带知识库 MCP server。 | Ingest Chinese content into your personal knowledge base — image/video routing, subtitle-first transcription, and a KB MCP server.
★ 667+8Star change over the last 7 days - #45
Android Input Method Editor (IME) based on Whisper
★ 633+4Star change over the last 7 days - #46
Fast, private, local-first voice app for Apple Silicon Macs — dictation, file/media transcription, meeting recording, Transforms, and a public automation CLI. Free and open-source.
★ 630+14Star change over the last 7 days - #47
A React component to make correcting automated transcriptions of audio and video easier and faster. By BBC News Labs. - Work in progress
★ 621+1Star change over the last 7 days - #48
C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more
★ 614+11Star change over the last 7 days - #49
Open-source, agent-driven editing pipeline: create subtitles, motion graphics, slides, edit and render videos.
★ 610—Star change over the last 7 days - #50★ 564+0Star change over the last 7 days
- #51★ 545-1Star change over the last 7 days
- #52
On-device meeting capture, dictation, and vault-native knowledge layer for macOS. Nothing leaves your Mac.
★ 518+1Star change over the last 7 days - #53
open source audio and video transcription software
★ 515+0Star change over the last 7 days - #54
Automatically synchronize and translate subtitles, or create new ones by transcribing, using pre-trained DNNs, Forced Alignments and Transformers. https://subaligner.readthedocs.io/
★ 510+1Star change over the last 7 days