{"repo":"aza-ali/webdownloader","free":true,"listed":false,"github":"https://github.com/aza-ali/webdownloader","clone":"git clone https://github.com/aza-ali/webdownloader.git","description":"Terminal CLI to download complete web pages (HTML + assets) for offline viewing. Markdown export, robots.txt aware.","language":"JavaScript","stars":44,"topics":["cli","command-line-tool","html-downloader","javascript","markdown-converter","nodejs","npm-package","offline-reading","robots-txt","web-archiving"],"license":"MIT","category":"cli-tools","readme_excerpt":"Web Downloader CLI A terminal tool for archiving complete web pages (HTML, CSS, JavaScript, images, fonts) for offline viewing or markdown export. wget grabs the file. curl fetches the bytes. web-downloader captures the whole page : rewrites links to local paths, follows the asset graph, and ships you a working offline copy or a clean markdown export. Features - Downloads HTML, CSS, JavaScript, images, and other assets - Rewrites links to local references so pages work offline - Optional conversion to Markdown - Configurable depth for following links - robots.txt modes: obey, warn, or ignore - Follows redirects automatically - Respects crawl-delay directives - Programmatic API for use as a Node.js module Install Usage Options Examples Programmatic API When to use what You want to... Reach for --- --- Download a single file curl or wget Mirror a static site wget --mirror Save one page to read offline (with images, working CSS) web-downloader Save a page as Markdown for LLM ingestion web-downloader -m Crawl a JS-heavy site A real browser tool (Playwright, Puppeteer) License MIT Contributing PRs welcome. Open an issue first for anything substantial.","default_branch":null,"files":null,"tree":[],"storefront":"/r/aza-ali","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/aza-ali/webdownloader/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}