HuCOLA

Hungarian Corpus

A collection of 9,076 annotated sentences in Hungarian to evaluate linguistic acceptability and grammaticality

Hungarian Corpus of Linguistic Acceptability

GitHub

1 stars
2 watching
0 forks
last commit: about 2 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
nytud/hucopaA dataset and annotation scheme for Hungarian causal reasoning tasks.1
nytud/huluA collection of linguistic datasets and benchmarks for natural language understanding tasks8
nytud/nytk-nerkorA Hungarian language named entity annotated corpus containing 1 million tokens with morphological annotation layers and various source files.15
nytud/huwsA dataset of manually curated Hungarian sentences with ambiguous wordings that require world knowledge and reasoning for resolution.1
nytud/panmorphHarmonized tagset and annotation scheme for Hungarian morphological analysers4
nytud/husstA dataset of annotated sentences for training and evaluating sentiment analysis models in the Hungarian language.1
vadno/korkor_pilotA large annotated corpus of Hungarian text with various linguistic annotations, split into development and test datasets for natural language processing tasks.2
poltextlab/hunempoli_corpusA manually annotated corpus for training and testing machine learning models of Aspect Based Sentiment Analysis (ABSA) in Hungarian language.0
nytud/hunlp-gateA collection of Hungarian NLP tools integrated as GATE processing resources8
nytud/emtsvA text processing system designed to handle various tasks in Hungarian language processing using Python and TSV-based data exchange.28
nytud/hadifogoly-adatbazisAn attempt to transcribe Cyrillic text into Hungarian script for a large dataset of WWII prisoner-of-war records23
nytud/quntokenA C++ tokenizer that tokenizes Hungarian text14
nytud/machine-translationProvides machine translation models and a demo site for Hungarian language translations5
nytud/huwnliA dataset and toolset for Hungarian anaphora resolution in natural language inference tasks0
elte-dh/regenykorpuszA large corpus of Hungarian novels with annotated texts and metadata, developed by the Department of Digital Humanities at Eötvös Loránd University.4