{"repo":"brightdata/mls-scraper","free":true,"listed":false,"github":"https://github.com/brightdata/mls-scraper","clone":"git clone https://github.com/brightdata/mls-scraper.git","description":"Scrapers for MLS real estate data with escalating anti-bot strategies: from HTTP requests to managed browsers, bypassing CAPTCHAs and blocks.","language":"Python","stars":13,"topics":["browser-automation","captcha-solving","curl-cffi","mls","playwright","python","real-estate","web-scraping"],"license":null,"category":"scrapers-browser-automation","readme_excerpt":"MLS Real Estate Data Scraper This repository provides Python scrapers to extract public real estate data from MLS portals. Accessing MLS-affiliated sites is difficult due to robust anti-bot defenses, including IP rate limiting, browser fingerprinting, and CAPTCHAs. The project uses an escalating strategy, from an HTTP scraper to a managed browser solution, to bypass these defenses and scale data extraction. Table of Contents - Why this data matters - Installation - Part 1 – The first scrape - Part 2 – Bypassing IP and browser blocks - Part 3 – Why local Playwright fails - Part 4 – Solving JavaScript and CAPTCHAs - Choosing your path Why this data matters This public MLS data is the raw material for any serious real estate strategy. Teams use it to: - Analyze market trends. Spotting pricing shifts, days-on-market, and inventory levels. - Build investment models. Finding and evaluating foreclosure or new-build opportunities. - Run competitive intelligence. Understanding what other builders and agencies are listing in real-time. The targets - mls.foreclosure.com (for foreclosure listings) - newhomesource.com (for new community and home listings) Technical stack - Python 3.10+ - curl cffi – a Python client that impersonates browser TLS/JA3 fingerprints - BeautifulSoup 4 – for parsing static HTML - Playwright – for automating browser actions and handling JavaScript-rendered content - Bright Data – provides the unblocking infrastructure (residential proxies & Browser API) for scali","default_branch":null,"files":null,"tree":[],"storefront":"/r/brightdata","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/brightdata/mls-scraper/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}