uptrain

AI evaluation platform

An open-source platform to evaluate and improve Generative AI applications through automated checks and analysis

UpTrain is an open-source unified platform to evaluate and improve Generative AI applications. We provide grades for 20+ preconfigured checks (covering language, code, embedding use-cases), perform root cause analysis on failure cases and give insights on how to resolve them.

GitHub

2k stars
21 watching
193 forks
Language: Python
last commit: about 2 years ago
Linked from 1 awesome list

autoevaluationevaluationexperimentationhallucination-detectionjailbreak-detectionllm-evalllm-promptingllm-testllmopsmachine-learningmonitoringopenai-evalsprompt-engineeringroot-cause-analysis

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
upgini/upginiAutomated data search and enrichment tool for machine learning pipelines321
upb-lea/gym-electric-motorA Python toolbox for simulating and controlling electric motors with a focus on reinforcement learning and classical control.311
cloud-cv/evalaiA platform for comparing and evaluating AI and machine learning algorithms at scale1,779
openai/finetune-transformer-lmThis project provides code and model for improving language understanding through generative pre-training using a transformer-based architecture.2,167
qiangyt/batchaiAutomates bulk code checks and generates unit test codes to supplement AI tools like Copilot and Cursor36
packtpublishing/hands-on-intelligent-agents-with-openai-gymTeaching software developers to build intelligent agents using deep reinforcement learning and OpenAI Gym374
uber-research/upsnetDevelops an instance segmentation and panoptic segmentation model for computer vision tasks.648
ethicalml/xaiAn eXplainability toolbox for machine learning that enables data analysis and model evaluation to mitigate biases and improve performance1,135
shu223/ios-genai-samplerA collection of Generative AI examples on iOS80
promptslab/openai-detectorAn AI classifier designed to determine whether text is written by humans or machines.122
codeintegrity-ai/mutahunterAutomated unit test generation and mutation testing tool using Large Language Models.252
ukgovernmentbeis/inspect_aiA framework for evaluating large language models669
oeg-upm/gtfs-benchProvides a benchmarking framework for evaluating declarative knowledge graph construction engines in the transport domain17
aporia-ai/mlnotifyAutomated notification system for machine learning model training343
trypromptly/llmstackA tool for building and deploying generative AI applications with a no-code multi-agent framework1,659