{"repo":"Fanfulla/OCR-buddy","free":true,"listed":false,"github":"https://github.com/Fanfulla/OCR-buddy","clone":"git clone https://github.com/Fanfulla/OCR-buddy.git","description":"Faithful, 100% local OCR Chrome extension | code, prose, formulas (LaTeX) & tables. No server, nothing leaves your device, no hallucinated text. PP-OCRv5 on ONNX Runtime Web.","language":"TypeScript","stars":48,"topics":["chrome-extension","ocr","formula-recognition","in-browser","latex-ocr","local-first","manifest-v3","offline","onnx","onnxruntime-web"],"license":null,"category":"media-processing","readme_excerpt":"OCR Buddy Faithful, fully-local OCR for Chrome. Grab text from anything on screen — a region, the viewport, or a whole scrolling page — code in a paused video, a paragraph in a PDF, a formula, a table. Or turn an entire page into clean Markdown for an LLM. No server. No image ever leaves your machine. No hallucinated text. 🌐 ocr-buddy.com · 🧩 Chrome extension (Manifest V3) · 🔓 Free & open source (MIT) · 🛡️ 100% local, privacy-first --- Demo Silent autoplay loops. ▶ Watch in High Quality Video: demo 1 · demo 2. --- Why this exists Modern OCR is dominated by large autoregressive vision-language models. They top the benchmarks — and they invent fluent, plausible, wrong text the moment the pixels get unclear. For most uses that's an annoyance. For code, numbers, prices, IDs, or anything you intend to trust , a confidently-wrong transcription is worse than no transcription at all. Those models are also far too heavy to run in a browser tab. OCR Buddy is built on the opposite bet: faithfulness over fluency, and the whole pipeline on your device. The interesting part is that those two goals don't fight — they point at the same engineering choices. The thesis: classic OCR, not generative OCR Hallucination in OCR is largely architectural . A generative model predicts the next likely token , so when the image is ambiguous it falls back on its language prior and writes something that reads well but isn't there. The classic OCR family — detection + CTC recognition — has no such prior","default_branch":null,"files":null,"tree":[],"storefront":"/r/Fanfulla","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/Fanfulla/OCR-buddy/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}