{"repo":"santhoshse7en/news-fetch","free":true,"listed":false,"github":"https://github.com/santhoshse7en/news-fetch","clone":"git clone https://github.com/santhoshse7en/news-fetch.git","description":"A Python Package which helps to scrape all news details from any news websites","language":"Python","stars":229,"topics":["newspaper3k","google-search-using-python","news","scraper","scraper-engine","news-details","python","news-website","felix","extracts"],"license":"MIT","category":"scrapers-browser-automation","readme_excerpt":"news-fetch Python news scraper & article extractor — extract title, text, authors, date, image, and publisher from any news URL. No API key. Confidence scores included. Fetch news. Know why it worked. --- Why news-fetch? A lightweight alternative to newspaper3k / newspaper4k / trafilatura wrappers — with its own extraction engine , confidence scores , and bulk + proxy support. Feature news-fetch --- :---: News article extraction (title, body, authors, date, image) ✅ Confidence scores + extraction provenance ✅ Bulk scraping ( fetch many / fetch iter / CLI JSONL) ✅ Proxy + proxy rotation for thousands of URLs ✅ RSS / sitemap article discovery ✅ Async ( pip install news-fetch[async] ) ✅ Optional browser render ( pip install news-fetch[browser] ) ✅ Disk cache + robots.txt respect ✅ No API key / no account ✅ Small deps ( lxml , requests , python-dateutil , cssselect ) ✅ --- Install Requirements: Python 3.10+ --- Quick start Single URL From HTML (no network) Bulk scraping + proxies Strict mode (production pipelines) Discovery (RSS / sitemaps) Async CLI Cache / robots / browser Custom strategy plugin --- Article fields url · canonical url · title · description · text · authors · published at · modified at · publisher · language · image · keywords · section · summary · word count · reading time minutes · page type · is article · confidence · extraction · sources --- Links - PyPI: https://pypi.org/project/news-fetch/ - Docs: https://santhoshse7en.github.io/newsfetch doc/ - GitHub: htt","default_branch":null,"files":null,"tree":[],"storefront":"/r/santhoshse7en","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/santhoshse7en/news-fetch/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}