evidently

Model monitor

An observability framework for evaluating and monitoring the performance of machine learning models and data pipelines

Evidently is ​​an open-source ML and LLM observability framework. Evaluate, test, and monitor any AI-powered system or data pipeline. From tabular data to Gen AI. 100+ metrics.

GitHub

6k stars
48 watching
607 forks
Language: Jupyter Notebook
last commit: almost 2 years ago
Linked from 8 awesome lists

data-driftdata-qualitydata-sciencedata-validationgenerative-aihacktoberfesthtml-reportjupyter-notebookllmllmopsmachine-learningmlopsmodel-monitoringpandas-dataframe

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
confident-ai/deepevalA framework for evaluating large language models4,003
explodinggradients/ragasA toolkit for evaluating and optimizing Large Language Model applications with objective metrics, test data generation, and seamless integrations.7,598
giskard-ai/giskardAutomates the detection of performance, bias, and security issues in AI applications4,125
openai/evalsA framework for evaluating large language models and systems, providing a registry of benchmarks.15,168
eleutherai/lm-evaluation-harnessProvides a unified framework to test generative language models on various evaluation tasks.7,200
instructor-ai/instructorA Python library that simplifies working with structured outputs from large language models8,551
relari-ai/continuous-evalProvides a comprehensive framework for evaluating Large Language Model (LLM) applications and pipelines with customizable metrics455
ianarawjo/chainforgeAn environment for battle-testing prompts to Large Language Models (LLMs) to evaluate response quality and performance.2,413
pair-code/litAn interactive tool for analyzing and understanding machine learning models3,500
cleanlab/cleanlabAutomates data quality checks and model training with AI-driven methods to improve machine learning performance9,820
christophm/interpretable-ml-bookA comprehensive resource for explaining the decisions and behavior of machine learning models.4,811
interpretml/interpretAn open-source package for explaining machine learning models and promoting transparency in AI decision-making6,324
h2oai/mli-resourcesProvides tools and techniques for interpreting machine learning models483
aiplanethub/beyondllmAn open-source toolkit for building and evaluating large language models267
psycoy/mixevalAn evaluation suite and dynamic data release platform for large language models230