laion-datasets

AI dataset collection

A repository containing a collection of large datasets used for training and testing AI models, specifically designed to improve image-text matching capabilities.

Description and pointers of laion datasets

GitHub

239 stars
6 watching
9 forks
Language: HTML
last commit: almost 4 years ago

Related projects:

RepositoryDescriptionStars
faceperceiver/laion-faceProvides pre-trained face detection and analysis models using large-scale image-text data281
aitutorials/datasetsA comprehensive collection of datasets from various AI-related sources worldwide.46
laion-ai/clip_benchmarkEvaluates and compares the performance of various CLIP-like models on different tasks and datasets.632
laion-ai/clapA library for learning audio embeddings from text and audio data using contrastive language-audio pretraining1,457
laion-ai/aesthetic-predictorPredicts aesthetic quality of images using CLIP model embeddings491
logpai/loghubProvides a collection of system log datasets for AI-driven analytics research.1,883
aisegmentcn/matting_human_datasetsA large dataset of human matting images and corresponding results for training person segmentation models.615
mirfan899/urduA collection of Urdu language datasets for various NLP tasks and applications71
niraj-lunavat/artificial-intelligenceA comprehensive resource for learning and exploring Artificial Intelligence (AI) concepts and applications1,667
karthikncode/nlp-datasetsA curated list of Natural Language Processing datasets used to train and evaluate NLP models.919
pratyushmaini/llm_dataset_inferenceDetects whether a given text sequence is part of the training data used to train a large language model.23
radi-cho/datasetgptA command-line interface to generate textual datasets with Large Language Models293
lemondan/humanparsing-datasetA collection of detailed pixel-wise annotations for fashion images used in human parsing research.213
poio-nlp/poio-corpusA collection of language resources extracted from publicly available sources.7
aiplanethub/beyondllmAn open-source toolkit for building and evaluating large language models267