FAITHSCORE

Hallucination detector

Evaluates answers generated by large vision-language models to assess hallucinations

GitHub

27 stars
2 watching
4 forks
Language: Python
last commit: almost 2 years ago

Related projects:

RepositoryDescriptionStars
amazon-science/refcheckerAutomates fine-grained hallucination detection in large language model outputs325
bradyfu/woodpeckerA method to correct hallucinations in multimodal large language models without requiring retraining617
damo-nlp-sg/vcdAn approach to reduce object hallucinations in large vision-language models by contrasting output distributions derived from original and distorted visual inputs222
junyangwang0410/haelmA framework for detecting hallucinations in large language models17
tianyi-lab/hallusionbenchAn image-context reasoning benchmark designed to challenge large vision-language models and help improve their accuracy259
rucaibox/popeAn evaluation framework for detecting object hallucinations in vision-language models187
yiyangzhou/lureAnalyzing and mitigating object hallucination in large vision-language models to improve their accuracy and reliability.136
x-plug/mplug-halowlEvaluates and mitigates hallucinations in multimodal large language models82
openkg-org/easydetectA framework to detect and mitigate hallucinations in multimodal large language models48
assafbk/mocha_codeA unified framework and benchmark for detecting and mitigating hallucinations in open-vocabulary image captioning models13
yfzhang114/llava-alignDebiasing techniques to minimize hallucinations in large visual language models75
openmoss/halluqaAn evaluation framework for assessing the performance of large language models on question-answering tasks with hallucination detection111
lalbj/paiImproves the performance of large language models by intervening in their internal workings to reduce hallucinations83
yuqifan1117/hallucidoctorThis project provides tools and frameworks to mitigate hallucinatory toxicity in visual instruction data, allowing researchers to fine-tune MLLM models on specific datasets.41
junyangwang0410/amberAn LLM-free benchmark suite for evaluating MLLMs' hallucination capabilities in various tasks and dimensions98