robustness_metrics

Model Robustness Tool

A toolset to evaluate the robustness of machine learning models

GitHub

466 stars
11 watching
33 forks
Language: Jupyter Notebook
last commit: about 2 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
hendrycks/robustnessEvaluates and benchmarks the robustness of deep learning models to various corruptions and perturbations in computer vision tasks.1,030
borealisai/advertorchA toolbox for researching and evaluating robustness against attacks on machine learning models1,311
robustbench/robustbenchA standardized benchmark for measuring the robustness of machine learning models against adversarial attacks682
google-research/deep_opeProvides benchmarking policies and datasets for offline reinforcement learning85
guanghelee/neurips19-certificates-of-robustnessProvides a framework for computing tight certificates of adversarial robustness for randomly smoothed classifiers.17
jmgirard/mreliabilityTools for calculating consistency of observer measurements in various contexts41
edisonleeeee/greatxA toolbox for graph reliability and robustness against noise, distribution shifts, and attacks.85
modeloriented/fairmodelsA tool for detecting bias in machine learning models and mitigating it using various techniques.86
google-research/rldsA toolkit for storing and manipulating episodic data in reinforcement learning and related tasks.302
google/ml-fairness-gymAn open-source tool for simulating the long-term impacts of machine learning-based decision systems on social environments314
freedomintelligence/mllm-benchEvaluates and compares the performance of multimodal large language models on various tasks56
benhamner/metricsProvides implementations of various supervised machine learning evaluation metrics in multiple programming languages.1,632
i-gallegos/fair-llm-benchmarkCompiles bias evaluation datasets and provides access to original data sources for large language models115
sail-sg/mmcbenchA benchmarking framework designed to evaluate the robustness of large multimodal models against common corruption scenarios27
cmawer/reproducible-modelA project demonstrating how to create a reproducible machine learning model using Python and version control86