llm_dataset_inference

Dataset checker

Detects whether a given text sequence is part of the training data used to train a large language model.

Official Repository for Dataset Inference for LLMs

GitHub

23 stars
1 watching
4 forks
Language: Jupyter Notebook
last commit: about 2 years ago

Related projects:

RepositoryDescriptionStars
iamgroot42/mimirA Python package for measuring memorization in Large Language Models.126
radi-cho/datasetgptA command-line interface to generate textual datasets with Large Language Models293
mlcommons/inferenceMeasures the performance of deep learning models in various deployment scenarios.1,256
pythainlp/prachathai-67kAn article classification dataset created from news articles scraped from Prachathai.com with multiple benchmark models for multi-label classification16
karthikncode/nlp-datasetsA curated list of Natural Language Processing datasets used to train and evaluate NLP models.919
jagilley/fact-checkerA tool for fact-checking LLM outputs with self-ask using prompt chaining289
snowflake-labs/snowflake-arcticA project providing optimized stacks for fine-tuning and inference of large language models, focusing on low-latency and high-throughput performance.525
at-aaims/forgePre-training large language models on scientific data for downstream applications12
i-gallegos/fair-llm-benchmarkCompiles bias evaluation datasets and provides access to original data sources for large language models115
mirfan899/urduA collection of Urdu language datasets for various NLP tasks and applications71
mengtingwan/goodreadsProvides code samples and notebooks to download, read, and analyze Goodreads datasets for research purposes.252
bmander/busbuzzardAnalyzes GPS data to infer probabilistic schedules from transit vehicle movements10
ymcui/cmrc2018A collection of data for evaluating Chinese machine reading comprehension systems419
lter/lterdatasamplerA collection of curated environmental datasets from US LTER sites, designed for teaching and training in data science.48
truera/trulensA tool to evaluate and track the performance of large language model (LLM) experiments2,233