Skip to main content
buildradar
Sign in
Topic · image-to-text

image-to-text

Tracked open-source repos tagged image-to-text, sorted by stars.

Repos
7
Total stars
11,043
Avg. stars
1,578
Share
0.00%

Topics that frequently appear alongside image-to-text on the same repo.

Recent risers

Repos created in the last 90 days, tagged image-to-text.

  • mac-ocr@privatenumber

    macOS CLI for OCR and searchable PDFs using Apple's Vision framework

    506
  • modlens@liustack

    The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | 全网最强 DeepSeek Harness 外挂视觉插件,为 DeepSeek、GLM 等纯文本模型外挂视觉能力,粘贴图片即得结构化 JSON 证据(OCR、版面、语义)。

    3,839+134Star change over the last 7 days
  • tesseract-ocr-for-php@thiagoalessio

    A wrapper to work with Tesseract OCR inside PHP.

    3,039-1Star change over the last 7 days
  • MORT@kmonkeyhead

    MORT 번역기 프로젝트 - Real-time game translator with OCR

    1,739+11Star change over the last 7 days
  • PaddleMIX@PaddlePaddle

    Paddle Multimodal Integration and eXploration, supporting mainstream multi-modal tasks, including end-to-end large-scale multi-modal pretrain models and diffusion model toolbox. Equipped with high performance and flexibility.

    723+1Star change over the last 7 days
  • RSTGameTranslation@thanhkeke97

    🎮 Real-time Game Translation Tool | OCR + AI Translation | Windows Gaming | Open Source

    641+1Star change over the last 7 days
  • Flame-Code-VLM@Flame-Code-VLM

    Flame is an open-source multimodal AI system designed to translate UI design mockups into high-quality React code. It leverages vision-language modeling, automated data synthesis, and structured training workflows to bridge the gap between design and front-end development.

    562+1Star change over the last 7 days
  • mac-ocr@privatenumber

    macOS CLI for OCR and searchable PDFs using Apple's Vision framework

    506Star change over the last 7 days
← Back to topics