FazBrowse GitHub Viewer | Trending |
URL:
| Home
Tools: [Download Repo ZIP]   [Original HTTPS Page]

ocr · GitHub Topics · GitHub

#

ocr

Here are 12,997 public repositories matching this topic...

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

  • Updated Jul 22, 2026
  • Python

Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.

  • Updated Aug 19, 2026
  • Python

Tesseract Open Source OCR Engine (main repository)

  • Updated Aug 17, 2026
  • C++

OCR software, free and offline. 开源、免费的离线OCR软件。支持截屏/批量导入图片,PDF文档识别,排除水印/页眉页脚,扫描/生成二维码。内置多国语言库。

  • Updated Nov 20, 2025
  • Python

A community-supported supercharged document management system: scan, index and archive all your documents

  • Updated Aug 19, 2026
  • Python

ShareX is a free and open-source application that enables users to capture or record any area of their screen with a single keystroke. It also supports uploading images, text, and various file types to a wide range of destinations.

  • Updated Aug 20, 2026
  • C#

Pure Javascript OCR for more than 100 Languages 📖🎉🖥

  • Updated May 17, 2026
  • JavaScript

OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched

  • Updated Aug 18, 2026
  • Python

Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and etc.

  • Updated Dec 5, 2025
  • Python

PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.

  • Updated Aug 18, 2026
  • Java

🌈一个跨平台的划词翻译和OCR软件 | A cross-platform software for text translation and recognition.

  • Updated Jul 4, 2026
  • JavaScript

pix2tex: Using a ViT to convert images of equations into LaTeX code.

  • Updated Jan 18, 2025
  • Python

Convert documents to structured data effortlessly. Unstructured is open-source ETL solution for transforming complex documents into clean, structured formats for language models. Visit our website to learn more about our enterprise grade Platform product for production grade workflows, partitioning, enrichments, chunking and embedding.

  • Updated Aug 19, 2026
  • HTML

Convert documentation websites, GitHub repositories, and PDFs into Claude AI skills with automatic conflict detection

  • Updated Aug 9, 2026
  • Python

带带弟弟 通用验证码识别OCR pypi版

  • Updated Mar 10, 2026
  • Python

一个简洁优雅的词典翻译 macOS App。开箱即用,支持离线 OCR 识别,支持有道词典,🍎 苹果系统词典,🍎 苹果系统翻译,OpenAI,Gemini,DeepL,Google,Bing,腾讯,百度,阿里,小牛,彩云和火山翻译。A concise and elegant Dictionary and Translator macOS App for looking up words and translating text.

  • Updated Aug 17, 2026
  • Swift

视觉小说翻译器 / Visual Novel Translator

  • Updated Aug 19, 2026
  • C++

超轻量级中文ocr,支持竖排文字识别, 支持ncnn、mnn、tnn推理 ( dbnet(1.8M) + crnn(2.5M) + anglenet(378KB)) 总模型仅4.7M

  • Updated May 18, 2026
  • C++

OCR & Document Extraction using vision models

  • Updated May 20, 2025
  • TypeScript

A fast, helpful, and open-source document parser

  • Updated Aug 19, 2026
  • Rust

Improve this page

Add a description, image, and links to the ocr topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the ocr topic, visit your repo's landing page and select "manage topics."

Learn more


Back | FazBrowse Home | New Git URL