{"repo":"hamzaskhan/HTTracker-Alternate","free":true,"listed":false,"github":"https://github.com/hamzaskhan/HTTracker-Alternate","clone":"git clone https://github.com/hamzaskhan/HTTracker-Alternate.git","description":"A python script that does what HTTrack does. Since this is in script form, you can also make Docker images and even improve. Please make it adaptable to mac I don't have a machine so can't really test it.","language":"Python","stars":11,"topics":["beautifulsoup4","crawler","docker","offline-browsing","python","requests-library-python","web-archiving","web-scraping","website-cloner","website-downloader"],"license":"MIT","category":"scrapers-browser-automation","readme_excerpt":"Picnic Day Series Offline Site Copier This Python script replicates the core functionality of HTTrack by crawling a website, downloading its pages, assets, and rewriting links for offline browsing. It creates a local copy of the site—including a JSON file representing the site's link tree and a CSV file of unique URLs—making it easy to browse offline. The script is designed to be easily adaptable across platforms, including macOS. Currently it works for windows but you are welcome to extend it's capability across MacOs directory structure. Features - Recursive Crawling: Follow internal links to a specified depth. - Asset Management: Downloads CSS, JavaScript, images, and rewrites paths to point to local copies. - Offline Saving: Saves the entire site structure under a designated directory. - Link Tree & Unique URLs: Generates a JSON file of the site structure and a CSV list of unique links. - Cross-Platform Compatibility: Adaptable for macOS and other environments. How to Use 1. Run the Script: Execute the script using Python 3. 2. Provide Inputs: - Enter the website URL to crawl (default: https://hamzak.cloud ). - Specify the maximum depth for crawling (default: 1 ). 3. View Outputs: - The offline site is saved under the downloaded site directory. - A JSON file ( link tree.json ) and a CSV file ( unique links.csv ) are created with the site's link data. 4. Enjoy Offline Browsing! Open the saved HTML files locally and browse the site without an internet connection. --- requir","default_branch":null,"files":null,"tree":[],"storefront":"/r/hamzaskhan","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/hamzaskhan/HTTracker-Alternate/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}