{"owner":"ExtractPDF4J","github":"https://github.com/ExtractPDF4J","claimed":false,"inventory":[],"indexed":[{"repo":"ExtractPDF4J/ExtractPDF4J","github":"https://github.com/ExtractPDF4J/ExtractPDF4J","description":"Java PDF table extraction & OCR library. Extract structured tables from text-based and scanned PDFs using stream, lattice (OpenCV-style grid detection), and hybrid parsing.","language":"Java","stars":547,"topics":["cli","document-processing","java","java17","maven","ocr","ocr-recognition","pdf-document","pdf-document-processor","pdf-extraction"],"license":null,"category":"media-processing"}],"how_to_buy":"GET /r/ExtractPDF4J/<repo> (Accept: application/json) for any listed repo here: tree, README, price and the checkout to pay (x402; rehearse first at its test twin, simulated money). Repos under 'indexed' are free: clone them from GitHub."}