ua-datasets

Ukrainian NLP datasets

Provides a collection of datasets for natural language processing in Ukrainian.

A collection of datasets for Ukrainian language

GitHub

57 stars
3 watching
2 forks
Language: Python
last commit: about 2 years ago
Linked from 1 awesome list

datasetnatural-language-processingnlpnlp-datasetsquestion-answeringtext-classificationtoken-classificationukrainian-language

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
robinhad/krukA collection of Ukrainian language models and datasets for natural language processing tasks.86
helsinki-nlp/ukrainianltA collection of Ukrainian language tools and resources for machine translation, natural language processing, and text translation.30
mirfan899/urduA collection of Urdu language datasets for various NLP tasks and applications71
karthikncode/nlp-datasetsA curated list of Natural Language Processing datasets used to train and evaluate NLP models.919
lang-uk/ner-ukA Ukrainian NER corpus and annotation dataset for training and evaluating named entity recognition models.90
grammarly/ua-gecA collection of annotated data and tools for improving the grammar and fluency of Ukrainian texts.255
alexa/massiveA collection of tools and modeling code for a large multilingual Natural Language Understanding dataset541
poio-nlp/poio-corpusA collection of language resources extracted from publicly available sources.7
amakukha/stemmers_ukrainianA novel stemmer for the Ukrainian language trained with AI28
universaldependencies/ud_ukrainian-iuA dataset of annotated text in Ukrainian with standardized formatting and annotation guidelines.27
piskvorky/gensim-dataA repository of pre-trained NLP models and corpora for text processing.990
pkuchmiichuk/ua-corefA dataset and tools for coreference resolution in Ukrainian language using OntoNotes 5.0 data and machine translation models.7
sandeep42/anuvadaThis is an open source PyTorch library providing tools and models to explain the predictions of deep neural networks for natural language processing tasks.19
nytud/huluA collection of linguistic datasets and benchmarks for natural language understanding tasks8
sdadas/polish-nlp-resourcesPre-trained models and resources for Natural Language Processing in Polish329