katana

Crawler

A fast and configurable web crawling framework that can be used to automate tasks in pipelines.

A next-generation crawling and spidering framework.

GitHub

13k stars
94 watching
656 forks
Language: Go
last commit: almost 2 years ago
Linked from 1 awesome list

clicrawlergocrawlerheadlessspider-frameworkweb-spider

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
jaeles-project/gospiderA tool for web crawling and exploitation written in Go.2,598
yujiosaka/headless-chrome-crawlerA distributed crawling framework that leverages Headless Chrome to scrape dynamic websites5,534
azure/cloud-katanaAutomates security assessment and research in cloud-native environments using event-driven serverless computing250
spatie/crawlerA powerful web crawler written in PHP that can execute JavaScript and crawl multiple URLs concurrently.2,552
grafana/tankaA configuration framework for Kubernetes clusters2,429
shpota/goxygenAutomates the creation of web projects with Go and various front-end frameworks and databases.3,531
webdriverio/webdriverioA test automation framework for end-to-end and unit testing in browsers and mobile devices using WebDriver technology9,115
grafana/pyroscopeA platform that helps you identify and debug performance issues in your applications10,178
spinnaker/spinnakerA platform for managing software releases across multiple cloud environments in a safe and efficient manner.9,366
internetarchive/heritrix3A web crawler designed to collect and preserve digital artifacts while respecting site policies and load constraints.2,857
bda-research/node-crawlerA NodeJS-based web crawler and spider that extracts data from websites.6,718
kataras/irisA fast and feature-rich web framework for building scalable and efficient web applications in Go.25,289
gohugoio/hugoA fast and flexible tool for generating static websites with built-in support for various content formats.76,514
bowser-js/bowserA small JavaScript library used to detect and analyze browser properties.5,517
geziyor/geziyorA fast and flexible web crawling and scraping framework for extracting structured data from websites.2,646