{"repo":"AmadeusITGroup/CrawlerBox","free":true,"listed":false,"github":"https://github.com/AmadeusITGroup/CrawlerBox","clone":"git clone https://github.com/AmadeusITGroup/CrawlerBox.git","description":"CrawlerBox is an automated analysis framework designed for parsing emails and crawling embedded web resources.","language":"Python","stars":22,"topics":["cybersecurity","data-enrichment","email-analysis","phishing","security-research","security-tools","threat-analysis","url-analysis","web-crawler","qr-code-analysis"],"license":"Apache-2.0","category":"scrapers-browser-automation","readme_excerpt":"CrawlerBox Description CrawlerBox is an automated analysis framework designed for parsing emails and crawling embedded web resources. This infrastructure was developed to facilitate the study of evasive phishing emails reported by end users. For more detailed information on CrawlerBox , its functionality, and the results obtained, please refer to our paper \"A Closer Look at Modern Evasive Phishing Emails\". Figure 1: CrawlerBox Analysis Pipeline Getting started Installation CrawlerBox is meant to be run on Windows. Local installation Local installation can be done using uv Necessary dependencies and configuration First you need to install vcredist x64.exe from the Visual C++ Redistributable Packages for Visual Studio 2013. It is necessary for the working of the library responsible for reading QR codes (QReader). CrawlerBox relies on external services to operate (e.g., Cisco Umbrella and Shodan for data enrichment). Additionally, it connects to two external servers: one database for retrieving newly user-reported messages and another for storing the obtained results. Before running CrawlerBox , you must configure these dependencies. Please use the config.py file accordingly. Please also consider rewriting the functions in personalized config.py : fetch new emails by date , fetch new emails by id , and url rewrite . The two first functions should match your implemetation for fetching newly reported emails, and url rewrite is designed to extract and return a decoded URL from a gi","default_branch":null,"files":null,"tree":[],"storefront":"/r/AmadeusITGroup","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/AmadeusITGroup/CrawlerBox/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}