DOGA_scraper

Document scraper

A tool that extracts and converts Galician Official journal documents to different formats based on input year.

Galician Official journal scraper

GitHub

0 stars
3 watching
0 forks
Language: Ruby
last commit: about 12 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
jjelosua/parlamentogaliciaExtracts text from Galician Parlament session transcripts and conversations0
jaimeiniesta/metainspectorA Ruby gem for web scraping and extracting metadata from web pages.1,038
jakopako/goskyrA tool to simplify web scraping of list-like structured data from web pages36
miyagawa/web-scraperA Perl toolkit for extracting structured data from HTML documents using a DSL-like interface.104
meilisearch/docs-scraperAutomates scraping and indexing of documentation content into a search engine297
fcannizzaro/jsoup-annotationsA Java library that provides annotations to simplify HTML scraping and processing with Jsoup239
tjatse/node-readabilityAutomates web page scraping and text extraction to make any webpage readable343
felipecsl/wombatA Ruby-based web crawler and data extraction tool with an elegant DSL.1,315
davemolk/gogetjsTools for extracting and analyzing JavaScript files from web pages41
malfrats/xeuledocA tool to fetch information about public Google documents from various services856
benibela/xidelA tool to extract data from web pages using various query languages and selectors.690
gushonorato/mechanizeA web scraping and automation tool for Elixir.30
eureka101v/weibospidergoA tool for extracting data from Weibo social media platform using Go programming language and Colly library66
oscarotero/embedA PHP library to retrieve metadata and embed code from any web page2,100
yhat/scrapeA collection of utility functions and tools to simplify web scraping in Go.1,513