Skip to main content
buildradar
Sign in
Topic · whisper

whisper

Tracked open-source repos tagged whisper, sorted by stars.

125 repos
  • mlx-tune@ARahim3

    Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unsloth-compatible API.

    1,398+8Star change over the last 7 days
  • Multi-backend whisper app. Blazing fast. Mac-arm optimized. Easy install. Input a local file or url and this service will transcribe it using Whisper AI. Completely private and Free 🤯🤯🤯

    1,394+0Star change over the last 7 days
  • CrisperWhisper@nyrahealth

    Controllable Transcription. Verbatim ( every, filler, pause, stutter, vocal sound) , or intended ( what the speaker meant to say, optimized for readability) with word-level timestamps.

    1,371+13Star change over the last 7 days
  • voicemode@mbailey

    Natural voice conversations with Claude Code

    1,347+6Star change over the last 7 days
  • whisper-ctranslate2@Softcatala

    Whisper command line client compatible with original OpenAI client based on CTranslate2.

    1,346+2Star change over the last 7 days
  • voxtype@peteonrails

    Voice-to-text with push-to-talk for Wayland compositors

    1,338+42Star change over the last 7 days
  • gp.nvim@Robitx

    Gp.nvim (GPT prompt) Neovim AI plugin: ChatGPT sessions & Instructable text/code operations & Speech to text [OpenAI, Ollama, Anthropic, ..]

    1,320+0Star change over the last 7 days
  • MTools@HG-ha

    MTools 是一个功能强大的多功能桌面应用程序,集成了音视频处理、图片编辑、文本操作和编码工具,内置AI增强功能。旨在简化您的工作流程,提升生产效率

    1,305+5Star change over the last 7 days
  • claude-video-vision@jordanrendric

    Give Claude the ability to watch and understand videos — Claude Code plugin with frame extraction and multimodal audio analysis

    1,281+6Star change over the last 7 days
  • whisper@graphite-project

    Whisper is a file-based time-series database format for Graphite.

    1,262+0Star change over the last 7 days
  • kubeai@kubeai-project

    AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-text.

    1,256+2Star change over the last 7 days
  • Whisper-Finetune@yeyupiaoling

    Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inference and support Web deployment, Windows desktop deployment, and Android deployment

    1,222-1Star change over the last 7 days
  • truss@basetenlabs

    The simplest way to serve AI/ML models in production

    1,199+3Star change over the last 7 days
  • hyprwhspr@goodroot

    Native speech-to-text for Linux - Fast, accurate, private, and hackable system-wide dictation

    1,189+10Star change over the last 7 days
  • 视频音频生成字幕,生成srt文件。无需申请第三方API,本地实现音频转文本。基于Transformer的视频字幕生成框架。A GUI tool for generating subtitle from videos and generating srt files.

    1,179+3Star change over the last 7 days
  • biniou@Woolverine94

    a self-hosted webui for 30+ generative ai

    1,150+1Star change over the last 7 days
  • AI-Waifu-Vtuber@ardha27

    AI Vtuber for Streaming on Youtube/Twitch

    1,115+1Star change over the last 7 days
  • Whisperboard@Saik0s

    The open-source iOS app that's making quality voice transcription more accessible on mobile devices.

    1,102+2Star change over the last 7 days
  • whisper-writer@savbell

    💬📝 A small dictation app using OpenAI's Whisper speech recognition model.

    1,100+2Star change over the last 7 days
  • voquill@voquill

    Open source voice dictation technology

    1,006-1Star change over the last 7 days
  • whisper-flow@dimastatz

    Whisper-Flow is a framework designed to enable real-time transcription of audio content using OpenAI’s Whisper model. Rather than processing entire files after upload (“batch mode”), Whisper-Flow accepts a continuous stream of audio chunks and produces incremental transcripts immediately.

    958+21Star change over the last 7 days
  • transcriptionstream@transcriptionstream

    turnkey self-hosted offline transcription and diarization service with llm summary

    948+2Star change over the last 7 days
  • whisper.api@innovatorved

    This project provides an API with user level access support to transcribe speech to text using a finetuned and processed Whisper ASR model.

    914+0Star change over the last 7 days
  • OpenCluely@TechyCSR

    OpenCluely is a free, open source Cluely (alternative), built for technical interviews like DSA, OAs, and CP. It offers an invisible overlay, real-time AI help, Smart Image Processing for question capture, and multi-language support : 100% customizable and private.

    902+33Star change over the last 7 days
  • TwitchLib@TwitchLib

    C# Twitch Chat, Whisper, API and PubSub Library. Allows for chatting, whispering, stream event subscription and channel/account modification. Supports everything that supports .NETStandard 2.0

    891+3Star change over the last 7 days
  • VideoSubtitleGenerator@buxuku

    批量为本地视频生成字幕文件,并可将字幕文件翻译成其它语言, 跨平台支持 window, mac 系统

    873+2Star change over the last 7 days
  • subvert@aschmelyun

    Generate subtitles, summaries, and chapters from videos in seconds

    868+0Star change over the last 7 days
  • OpenAI-Unity@srcnalt

    An unofficial OpenAI Unity Package that aims to help you use OpenAI API directly in Unity Game engine.

    843+1Star change over the last 7 days
  • whisper-playground@saharmor

    Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/

    835+0Star change over the last 7 days
  • go-carbon@go-graphite

    Golang implementation of Graphite/Carbon server with classic architecture: Agent -> Cache -> Persister

    829+0Star change over the last 7 days
← Back to topics