{"repo":"proxidize/reddit-scraper","free":true,"listed":false,"github":"https://github.com/proxidize/reddit-scraper","clone":"git clone https://github.com/proxidize/reddit-scraper.git","description":"A Python Reddit scraper with dual-mode architecture: simple requests for small jobs, async + proxy rotation for large-scale scraping. Features captcha solving, rich CLI, and smart job-size detection.","language":"Python","stars":21,"topics":["python3","reddit","reddit-api","reddit-scraper","scraper","uv"],"license":"MIT","category":"scrapers-browser-automation","readme_excerpt":"Reddit Scraper Features Multiple Scraping Methods - JSON Endpoint Scraper - Fast scraping using Reddit's .json endpoints (no authentication required) - Advanced Requests Scraper - Custom pagination and bulk scraping capabilities Advanced Capabilities - Proxy Rotation - Automatic proxy switching with health monitoring - Captcha Solving - Automated captcha handling using Capsolver API - User Agent Rotation - Realistic browser simulation - Rate Limiting - Respectful request throttling - Rich CLI Interface - Beautiful command-line interface with progress bars - Multiple Export Formats - JSON and CSV output with full comment thread data Installation Using uv (Recommended) Using pip Development Setup for Development Running Tests Test Markers - unit - Fast unit tests - integration - Integration tests that may hit external APIs - slow - Slow tests that should be skipped in CI Docker Support Building and Running with Docker Quick Start 1. Interactive Mode (Recommended) 2. Direct Commands Note : If you've properly installed the package with pip install -e . , you can use reddit-scraper directly instead of python3 -m reddit scraper.cli Configuration The scraper uses a JSON configuration file to manage all settings including proxies, captcha solvers, and scraping preferences. Copy config.example.json to config.json and edit: Key Features - Multiple Proxies : Add multiple HTTP and SOCKS5 proxies for automatic rotation - Captcha Solving : Integrate with Capsolver for automated captcha han","default_branch":null,"files":null,"tree":[],"storefront":"/r/proxidize","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/proxidize/reddit-scraper/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}