{"repo":"Topdu/OpenOCR","free":true,"listed":false,"github":"https://github.com/Topdu/OpenOCR","clone":"git clone https://github.com/Topdu/OpenOCR.git","description":"OpenOCR: An Open-Source Toolkit for General-OCR Research and Applications, integrates a unified training and evaluation benchmark, commercial-grade OCR and Document Parsing systems, and faithful reproductions of the core implementations from a wide range of academic papers.","language":"Python","stars":1435,"topics":["chineseocr","ocr","ocr-pytorch","scene-text-detection","scene-text-recognition","document-analysis","document-parsing","document-processing"],"license":"Apache-2.0","category":"media-processing","readme_excerpt":"OpenOCR: An Open-Source Toolkit for General-OCR Research and Applications If you find this project useful, please give us a star🌟. English 简体中文 OpenOCR is an open-source toolkit developed by the OCR team from FVL Lab, Fudan University, under the guidance of Prof. Yu-Gang Jiang and Prof. Zhineng Chen. It focuses on 「General-OCR」 tasks, including Text Detection and Recognition, Formula and Table Recognition , as well as Document Parsing and Understanding . The toolkit integrates a unified training and evaluation benchmark, commercial-grade OCR and Document Parsing systems, and faithful reproductions of the core implementations from a wide range of academic papers. OpenOCR aims to build a comprehensive open-source ecosystem for General-OCR, bridging academic research and real-world applications, and fostering the collaborative development and widespread deployment of OCR technologies across both research frontiers and industrial scenarios. We welcome researchers, developers, and industry partners to explore the toolkit and share feedback. 🚀 Quick Start Features - 🔥 OpenDoc-0.1B: Ultra-Lightweight Document Parsing System with 0.1B Parameters - ⚡\\[Quick Start\\] \\[Local Demo\\] - An ultra-lightweight document parsing system with only 0.1B parameters. - Two-stage pipeline: 1. Layout analysis via PP-DocLayoutV2 . 2. Unified recognition of text, formulas, and tables using the in-house model UniRec-0.1B - In the original version of UniRec-0.1B , only text and formula recognition were","default_branch":null,"files":null,"tree":[],"storefront":"/r/Topdu","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/Topdu/OpenOCR/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}