InfinityCrawler

Crawler library

A web crawling library for .NET that allows customizable crawling and throttling of websites.

A simple but powerful web crawler library for .NET

GitHub

248 stars
11 watching
36 forks
Language: C#
last commit: almost 3 years ago
Linked from 3 awesome lists

crawlerrobots-txtspiderweb-crawlerweb-crawling

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
brendonboshell/supercrawlerA web crawler designed to crawl websites while obeying robots.txt rules, rate limits and concurrency limits, with customizable content handlers for parsing and processing crawled pages.380
crypto-crawler/crypto-crawler-rsA Rust-based library for building and managing cryptocurrency crawlers235
cocrawler/cocrawlerA versatile web crawler built with modern tools and concurrency to handle various crawl tasks188
puerkitobio/gocrawlA concurrent web crawler written in Go that allows flexible and polite crawling of websites.2,036
fmpwizard/owlcrawlerA distributed web crawler that coordinates crawling tasks across multiple worker processes using a message bus.55
zhegexiaohuozi/seimicrawlerA distributed crawler framework that simplifies the process of building crawlers using Spring Boot and Redis1,980
postmodern/spidrA Ruby web crawling library that provides flexible and customizable methods to crawl websites809
feng19/spider_manA high-level web crawling and scraping framework for Elixir.23
hu17889/go_spiderA modular, concurrent web crawler framework written in Go.1,827
fredwu/crawlerA high-performance web crawling and scraping solution with customizable settings and worker pooling.945
webrecorder/browsertrix-crawlerA containerized browser-based crawler system for capturing web content in a high-fidelity and customizable manner.677
amoilanen/js-crawlerA Node.js module for crawling web sites and scraping their content254
wspl/creeperA framework for building cross-platform web crawlers using Go780
shapecrawler/shapecrawlerA .NET library for creating and manipulating PowerPoint presentations using Open XML.307
qinxuye/colaA high-level framework for building distributed data extractors from web pages1,501