{"repo":"scambier/obsidian-text-extractor","free":true,"listed":false,"github":"https://github.com/scambier/obsidian-text-extractor","clone":"git clone https://github.com/scambier/obsidian-text-extractor.git","description":"A (companion) plugin to facilitate the extraction of text from images (OCR) and PDFs.","language":"TypeScript","stars":626,"topics":["obsidian","obsidian-plugin","ocr","pdf"],"license":"GPL-3.0","category":"media-processing","readme_excerpt":"Obsidian Text Extractor --- ⚠️ Unmaintained ⚠️ I unfortunately can't dedicate any more time on Text Extractor. Pull requests are welcome, and you're free to fork and improve on this project. --- Text Extractor is a \"companion\" plugin. It's mainly useful when used in conjunction with other plugins (like Omnisearch), but you can also use it to quickly extract texts from images & PDFs . Supported files: - Images ( .png , .jpg , .jpeg , .webp , .gif , .bmp ) - PDFs ( .pdf ) - Office documents ( .docx , .xlsx ) Limitations - The plugin currently uses Tesseract.js and pdf-extract to extract texts from images and PDFs. Those libraries are not perfect, and may not work on some files. - 🟥 PDF files often fail to get their text extracted 🟥 . See #7 and #21 - 🟥 Text Extraction does not work on mobile 🟥 . Read the following section for more details. - Text Extractor needs an Internet connection to work. All the processing is done locally, but the language files needed by the underlying OCR library (Tesseract) are downloaded on demand. Cache & Sync The plugin caches the extracted texts as local small .json files inside the plugin directory. Those files can be synced between your devices. Since text extraction does not work on mobile, the plugin will use the synced cached texts if available. If not, an empty string will be returned. Installation Text Extractor is available on the Obsidian community plugins repository. You can also install it manually by downloading the latest release f","default_branch":null,"files":null,"tree":[],"storefront":"/r/scambier","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/scambier/obsidian-text-extractor/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}