{"repo":"triangle-motelti/web-scraper","free":true,"listed":false,"github":"https://github.com/triangle-motelti/web-scraper","clone":"git clone https://github.com/triangle-motelti/web-scraper.git","description":"Web Scraper is a compact Python tool for fetching web pages and extracting links, phone numbers, emails and addresses using requests and BeautifulSoup.","language":"Python","stars":17,"topics":["dns-resolver","http-requests","https","parser","python","python3","requests","requests-library-python","scraper","scraping"],"license":null,"category":"scrapers-browser-automation","readme_excerpt":"Triangle Web Scraper A lightweight, modular, and terminal-first Python web scraper for extracting links, phone numbers, emails (with optional MX verification), and addresses from web pages. Features a clean CLI interface, a startup ASCII banner, and graceful handling of Ctrl+C for quiet exits. Features - Modular Design : Extractors for links, phone numbers, emails, and addresses are organized in separate modules under extractors/ . - CLI-Friendly : Supports both interactive and non-interactive modes for flexible usage. - Graceful Exit : Exits cleanly with code 0 on Ctrl+C. - Optional MX Verification : Email extraction includes optional domain MX record checks (requires dnspython ). - Responsive : Uses requests with timeouts to handle network issues gracefully. - Customizable : Easily extendable with new extractors and configurable settings. Requirements The project includes a requirements.txt file with all necessary dependencies: - pyfiglet =0.8.1 — For ASCII banner generation. - colorama =0.4.6 — For colored terminal output. - requests =2.28.2 — For HTTP requests. - beautifulsoup4 =4.12.2 — For HTML parsing. - phonenumbers =8.13.12 — For phone number extraction and validation. - dnspython =2.4.2 (optional) — For MX record verification during email extraction. - Pillow =9.5.0 (optional) — For image processing features (if implemented). Notes : - dnspython is only required for email MX verification. - Pillow is only needed for image-related features (not used in core scraping)","default_branch":null,"files":null,"tree":[],"storefront":"/r/triangle-motelti","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/triangle-motelti/web-scraper/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}