COBS

Policy evaluation toolkit

A toolkit for evaluating and analyzing off-policy policy estimation methods in reinforcement learning

OPE Tools based on Empirical Study of Off Policy Policy Estimation paper.

GitHub

61 stars
3 watching
14 forks
Language: Python
last commit: about 4 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
zzhanghub/eval-co-sodAn evaluation tool for co-saliency detection tasks97
st-tech/zr-obpA framework for off-policy evaluation and learning in multi-armed bandit algorithms648
sony/pyieoeDevelops an interpretable evaluation procedure for off-policy evaluation (OPE) methods to quantify their sensitivity to hyper-parameter choices and/or evaluation policy choices.31
onlytailei/carla_cil_pytorchImplementation of a conditional imitation learning policy in PyTorch for autonomous driving using the Carla dataset.65
lartpang/pysodevaltoolkitA comprehensive Python toolbox for evaluating salient object detection and camouflaged object detection tasks168
cbfinn/gpsAn implementation of guided policy search and LQG-based trajectory optimization for reinforcement learning599
psecio/propauthEvaluates policies against user credentials and properties to determine access permissions.59
albermax/innvestigateA toolbox to help understand neural networks' predictions by providing different analysis methods and a common interface.1,271
cmlplatform/pycirkSoftware to model Circular Economy policy and technological interventions in Environmental Extended Input-Output Analysis20
hkust-nlp/cevalAn evaluation suite providing multiple-choice questions for foundation models in various disciplines, with tools for assessing model performance.1,650
nci/scoresA collection of tools and functions for evaluating and optimizing forecasts and models in various scientific fields.85
gzcch/bingoAn analysis project investigating limitations of visual language models in understanding and processing images with potential biases and interference challenges.53
mhubii/ppo_libtorchAn implementation of the proximal policy optimization algorithm in PyTorch.73
princeton-nlp/charxivAn evaluation suite for assessing chart understanding in multimodal large language models.85
krrishdholakia/betterpromptAn API for evaluating the quality of text prompts used in Large Language Models (LLMs) based on perplexity estimation43