Awesome Lists

chinese-llm-benchmark

by jeinlee1991

pushed almost 2 years ago

中文大模型能力评测榜单:目前已囊括128个大模型,覆盖chatgpt、gpt-4o、谷歌gemini、百度文心一言、阿里通义千问、百川、讯飞星火、商汤senseChat、minimax等商用模型, 以及qwen2.5、llama3.1、glm4、书生internLM2.5、openbuddy、AquilaChat等开源大模型。不仅提供能力评分排行榜,也提供所有模型的原始输出结果!

AI summary

LLM Benchmark

A comprehensive benchmarking platform for large language models, evaluating their performance across various capabilities and providing rankings and detailed results.

stars
3.1K
forks
137
watching
38

Add a GitHub project

Missing a project or an awesome list? Paste its GitHub URL and we fetch it right away.