CHAOS-evaluation

Segmentation evaluator

Evaluates segmentation performance in medical imaging using multiple metrics

Evaluation code of CHAOS challenge in MATLAB, Python and Julia languages.

GitHub

57 stars
1 watching
7 forks
Language: MATLAB
last commit: about 7 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
martinkersner/py-img-seg-evalA Python package providing metrics and tools for evaluating image segmentation models282
zhourax/vegaDevelops a multimodal task and dataset to assess vision-language models' ability to handle interleaved image-text inputs.33
ruixiangcui/agievalEvaluates foundation models on human-centric tasks with diverse exams and question types714
mfaruqui/eval-word-vectorsA set of Python scripts for evaluating word vectors on various tasks and comparing similarity between words.120
mshukor/evalign-iclEvaluating and improving large multimodal models through in-context learning21
krrishdholakia/betterpromptAn API for evaluating the quality of text prompts used in Large Language Models (LLMs) based on perplexity estimation43
aka-discover/ccmba_cvpr23Improving semantic segmentation robustness to motion blur using custom data augmentation techniques6
chenllliang/mmevalproA benchmarking framework for evaluating Large Multimodal Models by providing rigorous metrics and an efficient evaluation pipeline.22
benhamner/metricsProvides implementations of various supervised machine learning evaluation metrics in multiple programming languages.1,632
maluuba/nlg-evalA toolset for evaluating and comparing natural language generation models1,350
princeton-nlp/charxivAn evaluation suite for assessing chart understanding in multimodal large language models.85
uzh-rpg/rpg_trajectory_evaluationA toolbox for evaluating trajectory estimates in visual-inertial odometry, providing common methods and error metrics.1,076
freedomintelligence/mllm-benchEvaluates and compares the performance of multimodal large language models on various tasks56
sony/pyieoeDevelops an interpretable evaluation procedure for off-policy evaluation (OPE) methods to quantify their sensitivity to hyper-parameter choices and/or evaluation policy choices.31
pkinney/segseg_exAn Elixir module that calculates intersection and classification of two line segments6