AGLA
by Lackel
Pythonpushed about 2 years ago
[Arxiv 2024] AGLA: Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention
AI summary
Image descriptor model
Improves large vision-language models' ability to accurately describe images by combining global and local attention mechanisms.
- stars
- 18
- forks
- 0
- watching
- 2