PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.
Ocr Recognition GitHub Repositories
Explore popular GitHub repositories tagged “ocr-recognition”.
Compare stars, forks, and programming language using the same GitStar view as GitHub Trending.
Trending Repositories
🖼️ Image Toolbox is a powerful app for advanced image manipulation. It offers dozens of features, from basic tools like crop and draw to filters, OCR, and a wide range of image processing options
A fast, helpful, and open-source document parser
Self-hosted AI accounting app. LLM analyzer for receipts, invoices, transactions with custom prompts and categories
Text recognition (optical character recognition) with deep learning methods, ICCV 2019
开源易用的中文离线OCR,识别率媲美大厂,并且提供了易用的web页面及web的接口,方便人类日常工作使用或者其他程序来调用~
A curated list of resources for text detection/recognition (optical character recognition ) with deep learning methods.
Rust multi‑backend OCR/VLM engine (DeepSeek‑OCR-1/2, PaddleOCR‑VL, DotsOCR) with DSQ quantization and an OpenAI‑compatible server & CLI – run locally without Python.
Python tool for grabbing text via screenshot
A self-hosted file conversion server that supports 488 file formats in 26 languages.
A Tensorflow model for text recognition (CNN + seq2seq with visual attention) available as a Python package and compatible with Google Cloud ML Engine.
Convolutional Recurrent Neural Networks(CRNN) for Scene Text Recognition
No repository description provided.
🔥🔥🔥Java免费离线AI算法工具箱,支持人脸识别,活体检测,表情识别、目标检测、实例分割、行人检测、OCR文字识别、车牌识别、表格识别、ASR+TTS、机器翻译等功能,Maven引用即可使用。支持PyTorch、Tensorflow,已集成 Mtcnn、InsightFace、SeetaFace6、YOLOv8~v12、PaddleOCR(PPOCRv5)、Whisper等主流模型
A minimalist SOTA LaTeX OCR model with only 20M parameters, running in browser. Full training pipeline available for self-reproduction. | 超轻量SOTA LaTeX公式识别模型,仅20M参数量,可在浏览器中运行。训练全流程代码开源,以便自学复现。
🎮 Real-time Game Translation Tool | OCR + AI Translation | Windows Gaming | Open Source
Tesseract based OCR for android
Official Implementation of SynthTIGER (Synthetic Text Image Generator), ICDAR 2021
Java PDF table extraction & OCR library. Extract structured tables from text-based and scanned PDFs using stream, lattice (OpenCV-style grid detection), and hybrid parsing.
A scene text recognition toolbox based on PyTorch
Fast, offline OCR for Node.js & C++. PP-OCRv6 with Core ML / WebGPU hardware acceleration — recognize text in images with confidence scores & coordinates. npm: @arcships/light-ocr
第一届西安交通大学人工智能实践大赛(2018AI实践大赛--图片文字识别)第一名;仅采用densenet识别图中文字
A curated list of resources dedicated to table recognition
Genalog is an open source, cross-platform python package allowing generation of synthetic document images with custom degradations and text alignment capabilities.
Convolutional Recurrent Neural Network (CRNN) for image-based sequence recognition using Pytorch
Extracting Tables from Document Images using a Multi-stage Pipeline for Table Detection and Table Structure Recognition
Image to Text is a web tool to extract text from any image using OCR
:camera: Computer-Vision Demos
Handwritten LaTeX drawing website with real-time local OCR detection inference
Receipt parser application written in dart.