Photon

Crawler

A fast and flexible web crawler designed to gather information from the internet

Incredibly fast crawler designed for OSINT.

GitHub

11k stars
323 watching
2k forks
Language: Python
last commit: about 2 years ago
Linked from 2 awesome lists

crawlerinformation-gatheringosintpythonspider

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
unclecode/crawl4aiA web crawling tool designed to extract structured data from the web for use in AI applications18,541
gocolly/collyA framework for extracting structured data from websites in a fast and elegant way23,444
hakluke/hakrawlerA tool for automatically discovering and crawling web application endpoints and assets4,528
apify/crawleeA tool for building reliable web scraping and browser automation pipelines in Node.js.16,081
yujiosaka/headless-chrome-crawlerA distributed crawling framework that leverages Headless Chrome to scrape dynamic websites5,534
internetarchive/heritrix3A web crawler designed to collect and preserve digital artifacts while respecting site policies and load constraints.2,857
dedsecinside/torbotAn OSINT tool for exploring and analyzing dark web sites using Tor network3,016
matthewmueller/x-rayA flexible web scraping framework for extracting data from websites with customizable selectors and pagination support.5,883
jaeles-project/gospiderA tool for web crawling and exploitation written in Go.2,598
cobrateam/splinterA Python test framework for automating web applications using Selenium and other tools.2,726
spatie/crawlerA powerful web crawler written in PHP that can execute JavaScript and crawl multiple URLs concurrently.2,552
finic-ai/finicProvides cloud-hosted browsers for automation and scraping tasks to avoid detection by websites.2,311
nabla-c0d3/sslyzeAn SSL/TLS scanning tool and Python library for assessing server security configurations3,290
smicallef/spiderfootAutomates information gathering and analysis from various data sources to support threat intelligence and cybersecurity efforts13,364
geziyor/geziyorA fast and flexible web crawling and scraping framework for extracting structured data from websites.2,646