Skip to main content
buildradar
Sign in
Topic · tesseract

tesseract

Tracked open-source repos tagged tesseract, sorted by stars.

Repos
25
Total stars
213,081
Avg. stars
8,523
Share
0.01%

Topics that frequently appear alongside tesseract on the same repo.

Recent risers

Repos created in the last 90 days, tagged tesseract.

  • ocr-it@thiagotigaz

    Chrome extension: pin a screen region once, then hotkey your way through a paginated document. OCR runs 100% offline via bundled Tesseract.

    380
  • tesseract@tesseract-ocr

    Tesseract Open Source OCR Engine (main repository)

    76,310+61Star change over the last 7 days
  • tesseract.js@naptha

    Pure Javascript OCR for more than 100 Languages 📖🎉🖥

    38,688+11Star change over the last 7 days
  • OCRmyPDF@ocrmypdf

    OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched

    34,654+43Star change over the last 7 days
  • PyMuPDF@pymupdf

    PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.

    10,628+26Star change over the last 7 days
  • xberg@xberg-io

    Polyglot document intelligence with a Rust core: extract text, metadata, images, tables, and structured data from 106 formats across 140 file extensions, plus code intelligence for 371 languages. Fifteen bindings, with CLI, REST API, and MCP server.

    9,254+21Star change over the last 7 days
  • tessdata@tesseract-ocr

    Trained models with fast variant of the "best" LSTM models + legacy models

    7,647+4Star change over the last 7 days
  • TagUI@aisingapore

    Free RPA tool by AI Singapore

    6,327+2Star change over the last 7 days
  • RPA-Python@tebelorg

    Python package for doing RPA

    5,496+2Star change over the last 7 days
  • gosseract@otiai10

    Go package for OCR (Optical Character Recognition), by using Tesseract C++ library

    3,134+2Star change over the last 7 days
  • tesseract-ocr-for-php@thiagoalessio

    A wrapper to work with Tesseract OCR inside PHP.

    3,039-2Star change over the last 7 days
  • llm_aided_ocr@Dicklesworthstone

    Enhances Tesseract OCR output using LLMs (local or API) for error correction, smart chunking, and markdown formatting of scanned PDFs

    2,996+3Star change over the last 7 days
  • OSS-DocumentScanner@ossappscollective

    Document scanning app

    2,393+4Star change over the last 7 days
  • tesserocr@sirfz

    A Python wrapper for the tesseract-ocr API

    2,171+0Star change over the last 7 days
  • textshot@ianzhao

    Python tool for grabbing text via screenshot

    1,772-1Star change over the last 7 days
  • J.A.R.V.I.S@GauravSingh9356

    Personal Assistant built using python libraries. It does almost anything which includes sending emails, Optical Text Recognition, Dynamic News Reporting at any time with API integration, Todo list generator, Opens any website with just a voice command, Plays Music, Wikipedia searching, Dictionary with Intelligent Sensing i.e. auto spell checking, Weather Reporting i.e. temp, wind speed, humidity, YouTube searching, Google Map searching, Youtube Downloading, etc.

    1,341+5Star change over the last 7 days
  • Fork of tess-two rewritten from scratch to support latest version of Tesseract OCR.

    939+1Star change over the last 7 days
  • ccextractor@CCExtractor

    CCExtractor - Official version maintained by the core team

    901+2Star change over the last 7 days
  • rtesseract@dannnylo

    Ruby library for working with the Tesseract OCR.

    883+0Star change over the last 7 days
  • scribeocr@scribeocr

    Web interface for recognizing text, proofreading OCR, and creating fully-digitized documents.

    810+0Star change over the last 7 days
  • tesstrain@tesseract-ocr

    Train Tesseract LSTM with make

    723+0Star change over the last 7 days
  • BetterOCR@junhoyeo

    🔍 Better text detection by combining multiple OCR engines (EasyOCR, Tesseract, and Pororo) with 🧠 LLM.

    639+0Star change over the last 7 days
  • tessdata_fast@tesseract-ocr

    Fast integer versions of trained LSTM models

    610+0Star change over the last 7 days
  • android-ocr@SubhamTyagi

    Tesseract based OCR for android

    599+2Star change over the last 7 days
  • Tesseract OCR wrapper for React Native

    595+0Star change over the last 7 days
  • officeParser@harshankur

    A robust, strictly-typed Node.js and Browser library for parsing office files into a rich Abstract Syntax Tree (AST) and generating high-fidelity output in multiple formats. Parses: docx · pptx · xlsx · odt · odp · ods · pdf · rtf · csv · md · html. Generates: Markdown · HTML · CSV · RTF · PDF · Plain Text · RAG Chunks

    535+1Star change over the last 7 days
← Back to topics