php-spider

Web Crawler

A flexible PHP web crawler with configurable traversal algorithms and filters.

A configurable and extensible PHP web spider

GitHub

1k stars
87 watching
232 forks
Language: PHP
last commit: over 2 years ago
Linked from 2 awesome lists


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
brendonboshell/supercrawlerA web crawler designed to crawl websites while obeying robots.txt rules, rate limits and concurrency limits, with customizable content handlers for parsing and processing crawled pages.380
spider-rs/spiderA tool for web data extraction and processing using Rust1,234
hu17889/go_spiderA modular, concurrent web crawler framework written in Go.1,827
spider/spiderA flexible graph database abstraction for PHP23
rivermont/spidyA simple command-line web crawler that automatically extracts links from web pages and can be run in parallel for efficient crawling340
hightman/pspiderA parallel web crawler framework built using PHP and MySQLi266
stewartmckee/cobwebA flexible web crawler that can be used to extract data from websites in a scalable and efficient manner226
amoilanen/js-crawlerA Node.js module for crawling web sites and scraping their content254
3nock/spidersuiteA cross-platform web spider/crawler tool for analyzing and mapping attack surfaces614
manning23/mspiderA Python-based tool for web crawling and data collection from various websites348
crawlzone/crawlzoneA PHP framework for asynchronous internet crawling and web scraping78
postmodern/spidrA Ruby web crawling library that provides flexible and customizable methods to crawl websites809
fmpwizard/owlcrawlerA distributed web crawler that coordinates crawling tasks across multiple worker processes using a message bus.55
feng19/spider_manA high-level web crawling and scraping framework for Elixir.23
holgerd77/django-dynamic-scraperAn app that allows you to manage Scrapy spiders through a Django admin interface.1,155