{"repo":"ScholliYT/Broken-Links-Crawler-Action","free":true,"listed":false,"github":"https://github.com/ScholliYT/Broken-Links-Crawler-Action","clone":"git clone https://github.com/ScholliYT/Broken-Links-Crawler-Action.git","description":"GitHub Action to check a website for broken links","language":"Python","stars":31,"topics":["404","crawler","not-found","hacktoberfest"],"license":"MIT","category":"scrapers-browser-automation","readme_excerpt":"Broken-Links-Crawler-Action This action checks all links on a website. It will detect broken links i.e. links that return HTTP Code 403, 404... Check the logs to see which links are broken and consequently cause this action to fail. In addition to the actual broken URL, the navigation path to that location is also displayed. Based on this work: https://github.com/healeycodes/Broken-Link-Crawler Demo: https://github.com/ScholliYT/devportfolio/actions?query=workflow%3A%22Site+Reliability%22 Inputs website url Required The URL of the website to check. You can provide a comma separated list of URLs if you would like to crawl multiple websites. include url prefix Optional Comma separated list of URL prefixes to include. You may only want to crawl URLs that use \"https://\". exclude url prefix Optional Comma separated list of URL prefixes to exclude (default mailto:,tel:). Some sites do not respond properly to bots, and you might want to exclude those known sites to prevent a failed build. include url suffix Optional Comma separated list of URL suffixes to include. You may only want to crawl URLs that end with \".html\". exclude url suffix Optional Comma separated list of URL suffixes to exclude. You may want to skip images by using \".jpg,.gif,.png\". include url contained Optional Comma separated list of URL substrings to include. You may only want to crawl URLs that are hosted on your primary domain, so you could use \"mydomain.com\", which would include all of your subdomains, regardle","default_branch":null,"files":null,"tree":[],"storefront":"/r/ScholliYT","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/ScholliYT/Broken-Links-Crawler-Action/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}