{"repo":"Actomaton/ActoCrawler","free":true,"listed":false,"github":"https://github.com/Actomaton/ActoCrawler","clone":"git clone https://github.com/Actomaton/ActoCrawler.git","description":"🕸️ Swift Concurrency-powered crawler engine on top of Actomaton.","language":"Swift","stars":22,"topics":["crawler","swift"],"license":"MIT","category":"scrapers-browser-automation","readme_excerpt":"🕸️ ActoCrawler ActoCrawler is a Swift Concurrency-powered crawler engine on top of Actomaton, with flexible customizability to create various HTML scrapers, image scrapers, etc. Products - ActoCrawler : Core crawler engine with generic dependency injection and queueing logic. - ActoCrawlerNetworking : Networking helpers such as NetworkSession , Response , and withNetworkSession . - ActoCrawlerHTML : HTML scraping helpers such as htmlScraper backed by SwiftSoup. - ActoCrawlerPlaywrightPy : Native headless-browser adapter using playwright-python via PythonKit. - ActoCrawlerPlaywrightJS : WebAssembly headless-browser adapter using npm playwright via JavaScriptKit. Example - Examples Headless Browser Adapters Two Playwright integrations are available. - ActoCrawlerPlaywrightPy is the native macOS path and is demonstrated by Examples/HeadlessBrowserPyExample . - ActoCrawlerPlaywrightJS is the WebAssembly path and is demonstrated by Examples/HeadlessBrowserJSExample . The JavaScriptKit adapter expects the surrounding JavaScript runtime to install Playwright on globalThis. actoCrawlerPlaywright.playwright before the wasm module is instantiated. The example bootstrap in Examples/HeadlessBrowserJSExample/main.mjs does this automatically. Acknowledgements - mattsse/voyager License MIT","default_branch":null,"files":null,"tree":[],"storefront":"/r/Actomaton","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/Actomaton/ActoCrawler/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}