{"owner":"Murrough-Foley","github":"https://github.com/Murrough-Foley","claimed":false,"inventory":[],"indexed":[{"repo":"Murrough-Foley/rs-trafilatura","github":"https://github.com/Murrough-Foley/rs-trafilatura","description":"Fast, accurate web content extraction in Rust. ML page-type classification, per-type extraction, confidence scoring. F1=0.966 on ScrapingHub (#1), F1=0.859 across 2,008 annotated pages (1,497 development + 511 held-out test","language":"Rust","stars":50,"topics":["content-extraction","rust","search-engine-optimization","trafilatura","web-scraping","machine-learning","nlp"],"license":"Apache-2.0","category":"machine-learning"}],"how_to_buy":"GET /r/Murrough-Foley/<repo> (Accept: application/json) for any listed repo here: tree, README, price and the checkout to pay (x402; rehearse first at its test twin, simulated money). Repos under 'indexed' are free: clone them from GitHub."}