HtmlPageParser

HTML parser

An HTML parsing library that converts web pages to structured data and then generates Markdown content from that data

A generic HTML parser

GitHub

1 stars
1 watching
0 forks
Language: Python
last commit: over 3 years ago

Related projects:

RepositoryDescriptionStars
scrapy/scrapelyA pure-python library for extracting structured data from HTML pages.1,865
kovidgoyal/html5-parserA fast HTML parser written in C, optimized for performance.682
imangazaliev/didomA fast and simple HTML parser with support for CSS selectors and XPath expressions.2,202
html5lib/html5lib-pythonA standards-compliant Python library for parsing and serializing HTML documents and fragments.1,138
skevo18/pyeditorjsA Python package for parsing and rendering content from Editor.js JSON data in HTML format.19
bupt1987/html-parserA fast and efficient HTML parser for PHP.525
servo/html5everAn HTML parser designed to meet the standards of modern web browsers2,171
ndmitchell/tagsoupA Haskell library for parsing and extracting information from HTML/XML documents233
terrier989/universal_htmlA cross-platform Dart package for parsing and manipulating HTML, XML, and CSS documents across various platforms.0
iabudiab/htmlkitAn Objective-C framework for parsing and serializing HTML documents240
choru-k/react-native-html-parserA JavaScript library for parsing HTML and XML documents across multiple platforms, including React Native and Titanium.84
egonschiele/handsomesoupA Haskell library that simplifies HTML parsing by providing CSS selectors and attribute extraction functions.123
kennethreitz/requests-htmlA Pythonic HTML parsing library providing intuitive and asynchronous web scraping capabilities.304
lexborisov/myhtmlA fast HTML parsing library written in C1,657
rust-scraper/scraperA Rust library for parsing and querying HTML documents using CSS selectors.1,961