headless-chrome-crawler

by yujiosaka · Web

Distributed crawler powered by Headless Chrome

New0 ratings5,635 starsActive
WebDeveloper ToolJavaScriptDockerfile
↓ Download 1.8.0GitHub⚑ Report
headless-chrome-crawler — image 1

About this project

Headless Chrome Crawler API Examples Tips Code of Conduct Contributing Changelog Distributed crawler powered by Headless Chrome Features Crawlers based on simple requests to HTML files are generally fast. However, it sometimes ends up capturing empty bodies, especially when the websites are built on such modern frontend frameworks as AngularJS, React and Vue.js. Powered by Headless Chrome, the crawler provides simple APIs to crawl these dynamic websites with the following features: Distributed crawling Configure concurrency, delay and retry Support both depth-first search and breadth-first search algorithm Pluggable cache storages such as Redis Support CSV and JSON Lines for…

Technologies

JavaScriptDockerfilechromechromiumcrawler

Project health

Inactive for over a year
Last update3 years ago
Contributors7
Latest release1.8.0
Open issues & PRs33
LicenseMIT
On GitHubsince 2017

GitHub

5,635
stars
404
forks
7
contributors
33
open issues & PRs
JavaScript
language
3 years ago
last commit
View on GitHub ↗All releases ↗

Reviews

out of 5 · 0 ratings
★★★★★
0%
★★★★
0%
★★★
0%
★★
0%
0%
Sign in to write a review
No reviews yet
Be the first to review headless-chrome-crawler.

Built by

yujiosaka
Imported from GitHub · not yet claimed on GitPalace
View developer pageSign in with GitHub to claim

Maintain yujiosaka/headless-chrome-crawler? Claiming verifies admin access through your GitHub account and gives you control of this listing.

You might also like