headless-chrome-crawler
Distributed crawler powered by Headless Chrome
About this project
Headless Chrome Crawler API Examples Tips Code of Conduct Contributing Changelog Distributed crawler powered by Headless Chrome Features Crawlers based on simple requests to HTML files are generally fast. However, it sometimes ends up capturing empty bodies, especially when the websites are built on such modern frontend frameworks as AngularJS, React and Vue.js. Powered by Headless Chrome, the crawler provides simple APIs to crawl these dynamic websites with the following features: Distributed crawling Configure concurrency, delay and retry Support both depth-first search and breadth-first search algorithm Pluggable cache storages such as Redis Support CSV and JSON Lines for…
Technologies
Project health
GitHub
Reviews
Built by
Maintain yujiosaka/headless-chrome-crawler? Claiming verifies admin access through your GitHub account and gives you control of this listing.
