rageval

RAG evaluator

An evaluation tool for Retrieval-augmented Generation methods

Evaluation tools for Retrieval-augmented Generation (RAG) methods.

GitHub

141 stars
7 watching
11 forks
Language: Python
last commit: almost 2 years ago
Linked from 1 awesome list

evalutionllmrag

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
stanford-futuredata/aresA tool for automatically evaluating RAG models by generating synthetic data and fine-tuning classifiers499
amazon-science/ragcheckerA framework for evaluating and diagnosing retrieval-augmented generation systems630
paesslerag/gvalAn expression evaluation library for Go that supports arbitrary expressions and parameters758
whyhow-ai/rule-based-retrievalA Python package that enables the creation and management of Retrieval Augmented Generation applications with filtering capabilities.229
allenai/olmo-evalA framework for evaluating language models on NLP tasks326
maja42/govalA Go library for evaluating arbitrary arithmetic, string, and logic expressions with support for variables and custom functions.160
nullne/evaluatorAn expression evaluator library written in Go.41
huggingface/evaluateAn evaluation framework for machine learning models and datasets, providing standardized metrics and tools for comparing model performance.2,063
mlabonne/llm-autoevalA tool to automate the evaluation of large language models in Google Colab using various benchmarks and custom parameters.566
antonmedv/golang-expression-evaluation-comparisonA benchmarking repository comparing the performance of different expression evaluation packages in Go.48
huggingface/lightevalAn all-in-one toolkit for evaluating Large Language Models (LLMs) across multiple backends.879
thedevsaddam/govalidatorValidate golang request data with simple rules inspired by Laravel's request validation1,324
mshukor/evalign-iclEvaluating and improving large multimodal models through in-context learning21
rlancemartin/auto-evaluatorAn evaluation tool for question-answering systems using large language models and natural language processing techniques1,065
krrishdholakia/betterpromptAn API for evaluating the quality of text prompts used in Large Language Models (LLMs) based on perplexity estimation43