Skip to main content
buildradar
Sign in
Topic · ocr

ocr

Tracked open-source repos tagged ocr, sorted by stars.

186 repos
  • iOS-OCR-Server@riddleling

    An iOS OCR Server Using Apple’s Vision Framework

    2,029+19Star change over the last 7 days
  • gImageReader@manisandro

    A Gtk/Qt front-end to tesseract-ocr.

    1,989+4Star change over the last 7 days
  • TRex@amebalabs

    Copy any text on your screen, stop retyping.

    1,900+3Star change over the last 7 days
  • ocrs@robertknight

    Rust library and CLI tool for OCR (extracting text from images)

    1,879+3Star change over the last 7 days
  • OnnxOCR@jingsongliujing

    基于PaddleOCR重构,并且脱离PaddlePaddle深度学习训练框架的轻量级OCR,推理速度超快 —— A lightweight OCR system based on PaddleOCR, decoupled from the PaddlePaddle deep learning training framework, with ultra-fast inference speed.

    1,867+5Star change over the last 7 days
  • AdvancedLiterateMachinery@AlibabaResearch

    A collection of original, innovative ideas and algorithms towards Advanced Literate Machinery. This project is maintained by the OCR Team in the Language Technology Lab, Tongyi Lab, Alibaba Group.

    1,835+1Star change over the last 7 days
  • textshot@ianzhao

    Python tool for grabbing text via screenshot

    1,772-1Star change over the last 7 days
  • extractous@yobix-ai

    Fast and efficient unstructured data extraction. Written in Rust with bindings for many languages.

    1,772+0Star change over the last 7 days
  • tarsier@reworkd

    Vision utilities for web interaction agents 👀

    1,761+0Star change over the last 7 days
  • MORT@kmonkeyhead

    MORT 번역기 프로젝트 - Real-time game translator with OCR

    1,739+6Star change over the last 7 days
  • mokuro@kha-white

    Read Japanese manga inside browser with selectable text.

    1,720+8Star change over the last 7 days
  • react-native-executorch@software-mansion

    Declarative way to run AI models in React Native on device, powered by ExecuTorch.

    1,707+5Star change over the last 7 days
  • MixTeX multimodal LaTeX, ZhEn, and, Table OCR. It performs efficient CPU-based inference in a local offline on Windows.

    1,641+2Star change over the last 7 days
  • ExtractThinker@enoch3712

    ExtractThinker is a Document Intelligence library for LLMs, offering ORM-style interaction for flexible and powerful document workflows.

    1,596+1Star change over the last 7 days
  • yomitoku@kotaro-kinoshita

    YomiTokuはAIを活用した日本語文書解析エンジンを提供するPythonパッケージです。 Yomitoku is an AI-powered document image analysis package designed specifically for the Japanese language.

    1,583+3Star change over the last 7 days
  • rpaframework@robocorp

    Collection of open-source libraries and tools for Robotic Process Automation (RPA), designed to be used with both Robot Framework and Python

    1,555+0Star change over the last 7 days
  • PaddleOCR-json@hiroi-sora

    OCR离线图片文字识别命令行windows程序,以JSON字符串形式输出结果,方便别的程序调用。提供各种语言API。由 PaddleOCR C++ 编译。

    1,544+1Star change over the last 7 days
  • docstrange@NanoNets

    Extract and convert data from any document, images, pdfs, word doc, ppt or URL into multiple formats (Markdown, JSON, CSV, HTML) with intelligent structured data extraction and advanced OCR.

    1,534+1Star change over the last 7 days
  • documind@DocumindHQ

    Open-source platform for extracting structured data from documents using AI.

    1,520-1Star change over the last 7 days
  • keras-ocr@faustomorales

    A packaged and flexible version of the CRAFT text detector and Keras CRNN recognition model.

    1,473+0Star change over the last 7 days
  • OpenOCR@Topdu

    OpenOCR: An Open-Source Toolkit for General-OCR Research and Applications, integrates a unified training and evaluation benchmark, commercial-grade OCR and Document Parsing systems, and faithful reproductions of the core implementations from a wide range of academic papers.

    1,441+2Star change over the last 7 days
  • ai-hands-on@Ramakm

    A group of notebooks and other files which can help you learn AI from scratch.

    1,438+4Star change over the last 7 days
  • TextSnatcher@RajSolai

    How to Copy Text from Images ? Answer is TextSnatcher !. Perform OCR operations in seconds on Linux Desktop.

    1,387+0Star change over the last 7 days
  • tr@myhub

    Free Offline OCR 离线的中文文本检测+识别SDK

    1,381+0Star change over the last 7 days
  • HRConvert2@zelon88

    A self-hosted, respource aware file conversion server supporting 488 formats in 26 languages.

    1,368+4Star change over the last 7 days
  • LaTeX_OCR_PRO@LinXueyuanStdio

    :art: 数学公式识别增强版:中英文手写印刷公式、支持初级符号推导(数据结构基于 LaTeX 抽象语法树)Math Formula OCR Pro, supports handwrite, Chinese-mixed formulas and simple symbol reasoning (based on LaTeX AST).

    1,311+1Star change over the last 7 days
  • MouseTooltipTranslator@ttop32

    Mouseover Translate Any Language At Once - Chrome Extension: PDF Translator, EBOOK, EPUB, OCR, TTS, NETFLIX, YOUTUBE DUAL SUBTITLES, GOOGLE DOCS, AI, VIEWER, GMAIL, WRITING, IMAGE, DUAL SUBS, MANGA, HOVER, DICTIONARY, WEBTOON, EDGE, JAPANESE, ENGLISH

    1,305+4Star change over the last 7 days
  • Capso@lzhgus

    Open-source screenshot and screen recording for macOS. The free, native alternative to CleanShot X. Built with Swift 6.0 and SwiftUI.

    1,284+11Star change over the last 7 days
  • open-semantic-search@opensemanticsearch

    Open Source research tool to search, browse, analyze and explore large document collections by Semantic Search Engine and Open Source Text Mining & Text Analytics platform (Integrates ETL for document processing, OCR for images & PDF, named entity recognition for persons, organizations & locations, metadata management by thesaurus & ontologies, search user interface & search apps for fulltext search, faceted search & knowledge graph)

    1,206+1Star change over the last 7 days
  • PaddleOCR inference in PyTorch. Converted from [PaddleOCR](https://github.com/PaddlePaddle/PaddleOCR)

    1,205+0Star change over the last 7 days
← Back to topics