{"repo":"PKHarsimran/website-downloader","free":true,"listed":false,"github":"https://github.com/PKHarsimran/website-downloader","clone":"git clone https://github.com/PKHarsimran/website-downloader.git","description":"Website-downloader is a powerful and versatile Python script designed to download entire websites along with all their assets. This tool allows you to create a local copy of a website, including HTML pages, images, CSS, JavaScript files, and other resources. It is ideal for web archiving, offline browsing, and web development.","language":"Python","stars":176,"topics":["automation","beautifulsoup","data-mining","html","internet-tools","offline-browsing","open-source","python","python-scripts","requests"],"license":"MIT","category":"workflow-automation","readme_excerpt":"🌐 Website Downloader CLI Turn any website you're authorized to copy into a fast, browsable offline mirror — with one command. A modern, hackable alternative to wget --mirror and HTTrack — built in pure Python, without dragging in a heavy crawler framework. 📖 Read the Wiki — full CLI reference, cookbook, and troubleshooting · 📦 Install from PyPI --- Open example backup/index.html in your browser — the whole site works from disk: pages, styles, scripts, images, fonts, and media, all with links rewritten for offline browsing. ⚡ Quick Start The classic script entry point still works too: 🤔 Why Not Just wget or HTTrack? Those tools are great — until you hit a modern website. This project exists for the gap between \"one-liner that misses half the assets\" and \"write your own Scrapy project.\" website-downloader wget --mirror HTTrack Scrapy --- :-: :-: :-: :-: Modern assets: srcset , data-src , poster , CSS @import , JS asset strings ✅ partial partial build it yourself JavaScript rendering (React, Vue, Next.js) ✅ Playwright ❌ ❌ plugin Incremental re-mirroring ( ETag / Last-Modified ) ✅ timestamps only ✅ manual Cookies + custom headers for authorized portals ✅ ✅ ✅ ✅ Selective CDN mirroring with a domain allowlist ✅ ❌ partial manual Zip + WARC export ✅ WARC ✅ ❌ manual Windows-safe paths (long paths, reserved names, query hashing) ✅ ❌ partial manual Small, readable Python codebase you can extend ✅ ❌ (C) ❌ (C) framework ✨ Highlights - 🚀 Fast — parallel page fetching ( --page-threads ","default_branch":null,"files":null,"tree":[],"storefront":"/r/PKHarsimran","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/PKHarsimran/website-downloader/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}