DocumentUnderstanding
by bikash
Research papers and code on information extraction from image/pdf
AI summary
Info extractor
Research and development of tools and techniques for extracting information from images and PDFs using deep learning and graph neural networks.
- stars
- 96
- forks
- 11
- watching
- 7
Similar projects
Found by comparing what the projects do, not just their names.
Text extractor benchmark
Evaluates PDF extraction tools' ability to extract meaningful text from scientific articles
eyurtsev/kor1.6K
Text extractor
An open-source wrapper around LLMs to extract structured data from text
Information extractor
Extracts and stores information from various sources using AI models to generate answers.
Citation extractor
A Firefox extension that extracts website information to create citations in a specific citation format.
Image delineation resources
A curated list of resources and papers on 3D and 2D delineation techniques for image processing and computer vision applications.
Document extractor
A Python library for extracting information from unstructured documents using AI techniques and customizable pipelines.
Table extractor
A framework for extracting information from tables in scientific literature using a rule-based approach.
Relation extractor
Extracts binary relationships from English sentences at scale
Data extractor
A tool to extract local data storage of an Android application in one click.
PDF extractor
A CoffeeScript library for extracting text from PDF files and creating searchable documents with OCR capabilities
Extractor
A structured extraction library powered by AI models and TypeScript schema validation
PDF Extractor
A Quarkus-based microservice to extract text from PDF files
Metadata Extractor
A .NET library for extracting metadata from various image, video, and audio file formats.
PDF extractor
A tool to extract text from PDFs and add a searchable layer to them
Info extractor
An LLM-based tool for extracting structured information from text with ontology-based grounding.