kor

Text extractor

An open-source wrapper around LLMs to extract structured data from text

LLM(😽)

GitHub

2k stars
15 watching
91 forks
Language: Python
last commit: almost 2 years ago
information-extractionllmnatural-languagenatural-language-processingnatural-language-understanding

Related projects:

RepositoryDescriptionStars
geeks-of-data/knowledge-gptExtracts and stores information from various sources using AI models to generate answers.283
cognesy/instructor-phpA PHP library that simplifies the integration of Large Language Models into applications by providing structured data extraction and validation.230
karlicoss/kobuddyExtracts data from Kobo eReader databases for analysis and backup152
recrm/archivetoolsA collection of tools for extracting and analyzing data from web archives71
monarch-initiative/ontogptAn LLM-based tool for extracting structured information from text with ontology-based grounding.626
bikash/documentunderstandingResearch and development of tools and techniques for extracting information from images and PDFs using deep learning and graph neural networks.96
wse-research/loris-llm-generated-representations-of-sparql-queriesGenerates natural language representations of SPARQL queries for knowledge graphs3
quanteda/spacyrAn R wrapper around spaCy for natural language processing tasks251
nikolamilosevic86/tabinoutA framework for extracting information from tables in scientific literature using a rule-based approach.42
richardlitt/lrlDeveloping tools and scripts to extract data from low-resource languages, focusing on language processing and machine learning applications.2
koraykv/fexA Lua-based library for feature extraction in computer vision applications using the SIFT algorithm10
philipperemy/stanford-openie-pythonProvides a Python interface to extract structured relation triples from plain text using CoreNLP's open information extraction system.639
microsoft/unicoderThis repository provides pre-trained models and code for understanding and generation tasks in multiple languages.89
stephenbrannon/iocextractorExtracts and organizes Indicators of Compromise from unstructured text files into structured formats.135
lang-uk/ner-ukA Ukrainian NER corpus and annotation dataset for training and evaluating named entity recognition models.90