aws-pdf-textract-pipeline
by aeksco
TypeScriptpushed over 2 years ago
Data pipeline for crawling PDFs from the Web and transforming their contents into structured data using AWS textract. Built with AWS CDK + TypeScript
AI summary
PDF extractor
A data pipeline for extracting structured data from PDFs using AWS Textract and cloud-based services
- stars
- 164
- forks
- 18
- watching
- 3
- awesome list
- 1
Featured in 1 awesome list
Each link jumps to the spot where the list mentions aws-pdf-textract-pipeline.