PHPScraper

Web scraper

A web scraping utility for PHP that simplifies the process of extracting information from websites.

A universal web-util for PHP.

GitHub

544 stars
18 watching
75 forks
Language: PHP
last commit: over 2 years ago
Linked from 1 awesome list

beautifulsoupchromiumheadless-chromephpphp-crawlerphp-scraperphp-spiderphp-spiderspuppeteerpyppeteerscraperscrapingscraping-websitesscrapyweb-scraperweb-scraping

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
spider-rs/spiderA tool for web data extraction and processing using Rust1,234
the-markup/blacklight-collectorA tool for scraping website content and analyzing browser behavior205
jakopako/goskyrA tool to simplify web scraping of list-like structured data from web pages36
benibela/xidelA tool to extract data from web pages using various query languages and selectors.690
slotix/dataflowkitA framework for extracting structured data from web pages using CSS selectors.667
joseconstela/webparsyA Node.js library and CLI for scraping websites using Puppeteer and YAML definitions44
propublica/uptonA web scraping framework that simplifies the process by handling repetitive tasks and provides options for efficient data retrieval1,612
oscarotero/embedA PHP library to retrieve metadata and embed code from any web page2,100
postmodern/spidrA Ruby web crawling library that provides flexible and customizable methods to crawl websites809
fimad/scalpelA web scraping library providing a declarative interface on top of an HTML parsing library to extract data from HTML pages325
scrapy/scrapelyA pure-python library for extracting structured data from HTML pages.1,865
rivermont/spidyA simple command-line web crawler that automatically extracts links from web pages and can be run in parallel for efficient crawling340
skallwar/suckitA Rust-based web scraping tool that recursively visits and downloads websites to disk.750
miyagawa/web-scraperA Perl toolkit for extracting structured data from HTML documents using a DSL-like interface.104
zhuyingda/websterA framework for automating web scraping and crawling tasks using Node.js518