Metrics

Evaluation metrics library

Provides implementations of various supervised machine learning evaluation metrics in multiple programming languages.

Machine learning evaluation metrics, implemented in Python, R, Haskell, and MATLAB / Octave

GitHub

2k stars
87 watching
454 forks
Language: Python
last commit: over 3 years ago
Linked from 3 awesome lists


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
martinkersner/py-img-seg-evalA Python package providing metrics and tools for evaluating image segmentation models282
statisticianinstilettos/recmetricsA library providing evaluation metrics and diagnostic tools for recommender systems.571
enochkan/torch-metricsA collection of common machine learning evaluation metrics implemented in PyTorch110
astrazeneca/rexmexA library providing a comprehensive set of metrics and tools for evaluating recommender systems278
mop/bierThis project implements a deep metric learning framework using an adversarial auxiliary loss to improve robustness.39
scikit-learn-contrib/metric-learnA Python library providing efficient implementations of various supervised and weakly-supervised metric learning algorithms.1,402
freedomintelligence/mllm-benchEvaluates and compares the performance of multimodal large language models on various tasks56
lartpang/pysodmetricsA library providing an implementation of various metrics for object segmentation and saliency detection in computer vision.150
pascaldekloe/metricsProvides a simple and efficient way to track and expose performance metrics in Go applications.28
hashicorp/go-metricsA Golang library for exporting performance and runtime metrics to external systems.1,470
mshukor/evalign-iclEvaluating and improving large multimodal models through in-context learning21
benhamner/machinelearning.jlA Julia library providing a consistent API for common machine learning algorithms116
szilard/benchm-mlA benchmark for evaluating machine learning algorithms' performance on large datasets1,874
beberlei/metricsA simple metrics library that abstracts different data collection backends.317
i-gallegos/fair-llm-benchmarkCompiles bias evaluation datasets and provides access to original data sources for large language models115