{"repo":"lonexreb/site2cli","free":true,"listed":false,"github":"https://github.com/lonexreb/site2cli","clone":"git clone https://github.com/lonexreb/site2cli.git","description":"Turn any website into a CLI/API for AI agents","language":"Python","stars":60,"topics":["ai-agents","api-discovery","cli","developer-tools","mcp","playwright","pypi","python"],"license":"MIT","category":"ai-agents","readme_excerpt":"Turn any website into a CLI/API for AI agents. Discover APIs automatically. Extract structured data like Firecrawl — but local, free, and open-source. Free, open-source alternative to Firecrawl, Browserbase browser-to-api, and Browserbase browser-to-cli — runs 100% locally, no API keys, no per-page billing. --- The Problem AI agents interact with websites through browser automation, which is slow, expensive, and unreliable: Without site2cli With site2cli --- --- --- Speed 10-30s per action (browser) 95% for discovered APIs Setup Write custom Playwright scripts site2cli discover Output Screenshots, raw HTML Structured JSON, typed clients CLI Overview What's New in v0.7.0 — browser-to-api + for everyone - Offline trace replay — site2cli discover --replay trace.json rebuilds the spec without re-running the browser. Save real captures with --save-trace . - Four artifacts every run — OpenAPI 3.1 ( .yaml / .json ), Python client, zero-dep JavaScript ES-module client (Node 18+ / Deno / Bun / browser), and a dark-theme HTML coverage report with gap candidates. - chunk command — RAG-ready chunking (fixed / sentence / heading) for .md , .txt , and PDFs. - search command — DuckDuckGo search piped into --scrape , --extract , or --chunk in one go. - PDF parsing — pdf to text , pdf to markdown , pdf page count via the [rag] extra. - 559 tests (up from 500), all passing in --- Extract & Scrape — Open-Source Firecrawl Alternative site2cli includes a complete web extraction pipeline — no API ","default_branch":null,"files":null,"tree":[],"storefront":"/r/lonexreb","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/lonexreb/site2cli/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}