MMCU
MEASURING MASSIVE MULTITASK CHINESE UNDERSTANDING
AI summary
Chinese understanding benchmark
Measures the understanding of massive multitask Chinese datasets using large language models
- stars
- 87
- forks
- 12
- watching
- 2
Similar projects
Found by comparing what the projects do, not just their names.
Language Model
A high-performance language model designed to excel in tasks like natural language understanding, mathematical computation, and code generation
Chinese Text Model
An implementation of a large language model for Chinese text processing, focusing on MoE (Multi-Headed Attention) architecture and incorporating a vast vocabulary.
Chart model trainer
Develops a large-scale dataset and benchmark for training multimodal chart understanding models using large language models.
Questionnaire
An evaluation suite to assess language models' performance in multi-choice questions
Video benchmarking toolkit
Evaluates and benchmarks large language models' video understanding capabilities
Chinese language models
Provides pre-trained models for Chinese language tasks with improved performance and smaller model sizes compared to existing models.
Language Model
This project makes a large language model accessible for research and development
LLM benchmarker
A benchmarking framework for large language models
Model evaluator
Evaluates and compares the performance of multimodal large language models on various tasks
Chinese NLU/NGL toolkit
This project provides pre-trained models and tools for natural language understanding (NLU) and generation (NLG) tasks in Chinese.
LM Benchmark
A benchmark for evaluating large language models in multiple languages and formats
Chinese language model
Trains a large Chinese language model on massive data and provides a pre-trained model for downstream tasks
Chinese Language Model
Trains and evaluates a Chinese language model using adversarial training on a large corpus.
Language Model
Large-scale language model with improved performance on NLP tasks through distributed training and efficient data processing
OCR Benchmark
An evaluation benchmark for OCR capabilities in large multmodal models.