image-to-text
Tracked open-source repos tagged image-to-text, sorted by stars.
Related topics
Topics that frequently appear alongside image-to-text on the same repo.
Recent risers
Repos created in the last 90 days, tagged image-to-text.
- #1
The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | 全网最强 DeepSeek Harness 外挂视觉插件,为 DeepSeek、GLM 等纯文本模型外挂视觉能力,粘贴图片即得结构化 JSON 证据(OCR、版面、语义)。
★ 3,839+134Star change over the last 7 days - #2
A wrapper to work with Tesseract OCR inside PHP.
★ 3,039-1Star change over the last 7 days - #3★ 1,739+11Star change over the last 7 days
- #4
Paddle Multimodal Integration and eXploration, supporting mainstream multi-modal tasks, including end-to-end large-scale multi-modal pretrain models and diffusion model toolbox. Equipped with high performance and flexibility.
★ 723+1Star change over the last 7 days - #5
🎮 Real-time Game Translation Tool | OCR + AI Translation | Windows Gaming | Open Source
★ 641+1Star change over the last 7 days - #6
Flame is an open-source multimodal AI system designed to translate UI design mockups into high-quality React code. It leverages vision-language modeling, automated data synthesis, and structured training workflows to bridge the gap between design and front-end development.
★ 562+1Star change over the last 7 days - #7★ 506—Star change over the last 7 days