wombat
Web scraper library
A Ruby-based web crawler and data extraction tool with an elegant DSL.
Lightweight Ruby web crawler/scraper with an elegant DSL which extracts structured data from pages.
1k stars
51 watching
129 forks
Language: Ruby
last commit: over 2 years agoLinked from 3 awesome lists
crawlerdslrubyscraper
Related projects:
| Repository | Description | Stars |
|---|---|---|
| A Scala-based DSL for programmatically accessing and interacting with web pages | 149 | |
| A tool to extract data from web pages using various query languages and selectors. | 690 | |
| A Ruby web crawling library that provides flexible and customizable methods to crawl websites | 809 | |
| A Ruby gem for web scraping and extracting metadata from web pages. | 1,038 | |
| A Scala library providing a DSL for loading and extracting content from HTML pages | 717 | |
| Downloads and crawls web pages, allowing for the archiving of websites. | 556 | |
| A Perl toolkit for extracting structured data from HTML documents using a DSL-like interface. | 104 | |
| A Node.js library and CLI for scraping websites using Puppeteer and YAML definitions | 44 | |
| A command line tool and Python library for extracting data from various web sources. | 293 | |
| A PHP library to retrieve metadata and embed code from any web page | 2,100 | |
| A framework for extracting structured data from web pages using CSS selectors. | 667 | |
| A tool for web data extraction and processing using Rust | 1,234 | |
| A tool that extracts and converts Galician Official journal documents to different formats based on input year. | 0 | |
| A utility for systematically extracting URLs from web pages and printing them to the console. | 268 | |
| A web scraping library providing a declarative interface on top of an HTML parsing library to extract data from HTML pages | 325 |