{"repo":"povilasb/scrapy-html-storage","free":true,"listed":false,"github":"https://github.com/povilasb/scrapy-html-storage","clone":"git clone https://github.com/povilasb/scrapy-html-storage.git","description":"Scrapy downloader middleware that stores response HTMLs to disk.","language":"Python","stars":18,"topics":["scrapy","middleware","python"],"license":"MIT","category":"scrapers-browser-automation","readme_excerpt":"About A Scrapy downloader middleware that stores response HTMLs to disk. Usage Turn downloader on, e.g. specifying it in settings.py : None of responses by default are saved to disk. You must select for which requests the response HTMLs will be saved: The file path where HTML will be stored is resolved with spider method response html path . E.g.: Configuration HTML storage downloader middleware supports such options: gzip output (bool) - if True, HTML output will be stored in gzip format. Default is False. save html on status (list) - if not empty, sets list of response codes whitelisted for html saving. If list is empty or not provided, all response codes will be allowed for html saving. Sample:","default_branch":null,"files":null,"tree":[],"storefront":"/r/povilasb","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/povilasb/scrapy-html-storage/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}