llm_dataset_inference
Official Repository for Dataset Inference for LLMs
AI summary
Dataset checker
Detects whether a given text sequence is part of the training data used to train a large language model.
- stars
- 23
- forks
- 4
- watching
- 1
Similar projects
Found by comparing what the projects do, not just their names.
Model memorization analysis
A Python package for measuring memorization in Large Language Models.
LLM dataset generator
A command-line interface to generate textual datasets with Large Language Models
Model benchmarking suite
Measures the performance of deep learning models in various deployment scenarios.
News dataset
An article classification dataset created from news articles scraped from Prachathai.com with multiple benchmark models for multi-label classification
NLP datasets
A curated list of Natural Language Processing datasets used to train and evaluate NLP models.
LLM fact checker
A tool for fact-checking LLM outputs with self-ask using prompt chaining
LLM inference stack
A project providing optimized stacks for fine-tuning and inference of large language models, focusing on low-latency and high-throughput performance.
Scientific LLMs
Pre-training large language models on scientific data for downstream applications
Bias datasets
Compiles bias evaluation datasets and provides access to original data sources for large language models
NLP datasets
A collection of Urdu language datasets for various NLP tasks and applications
Dataset explorer
Provides code samples and notebooks to download, read, and analyze Goodreads datasets for research purposes.
Schedule inference tool
Analyzes GPS data to infer probabilistic schedules from transit vehicle movements
Reading dataset
A collection of data for evaluating Chinese machine reading comprehension systems
Dataset library
A collection of curated environmental datasets from US LTER sites, designed for teaching and training in data science.
truera/trulens2.2K
Performance tracker for LLMs
A tool to evaluate and track the performance of large language model (LLM) experiments