LawBench

Legal model evaluator

Evaluates the legal knowledge of large language models using a custom benchmarking framework.

Benchmarking Legal Knowledge of Large Language Models

GitHub

273 stars
7 watching
44 forks
Language: Python
last commit: almost 3 years ago
benchmarkchatgptlawllm

Related projects:

RepositoryDescriptionStars
open-compass/vlmevalkitAn evaluation toolkit for large vision-language models1,514
freedomintelligence/mllm-benchEvaluates and compares the performance of multimodal large language models on various tasks56
mlabonne/llm-autoevalA tool to automate the evaluation of large language models in Google Colab using various benchmarks and custom parameters.566
qcri/llmebenchA benchmarking framework for large language models81
liuhc0428/law-gptA Chinese law-focused conversational AI model designed to provide reliable and professional legal answers.1,072
andrewzhe/lawyer-llamaAn AI model trained on legal data to provide answers and explanations in Chinese law871
obss/juryA comprehensive toolkit for evaluating NLP experiments offering automated metrics and efficient computation.187
iclrandd/blackstoneDevelops an NLP pipeline and model for processing long-form legal text641
open-compass/mmbenchA collection of benchmarks to evaluate the multi-modal understanding capability of large vision language models.168
oeg-upm/lubm4obdaEvaluates Ontology-Based Data Access systems with inference and meta knowledge benchmarking4
openai/simple-evalsEvaluates language models using standardized benchmarks and prompting techniques.2,059
siat-nlp/hanfeiDevelops and trains a large-scale, parameterized model for legal question answering and text generation105
maluuba/nlg-evalA toolset for evaluating and comparing natural language generation models1,350
openlmlab/gaokao-benchAn evaluation framework using Chinese high school examination questions to assess large language model capabilities565
mlgroupjlu/llm-eval-surveyA repository of papers and resources for evaluating large language models.1,450