Skip to content

Latest commit

 

History

History
25 lines (17 loc) · 1.32 KB

File metadata and controls

25 lines (17 loc) · 1.32 KB

Year 2024

  • Visual Hallucinations of Multi-modal Large Language Models, Wen Huang, Hongbin Liu, Minxin Guo, Neil Zhenqiang Gong [Paper]

  • Uncertainty-Aware Evaluation for Vision-Language Models, Vasily Kostumov, Bulat Nutfullin, Oleg Pilipenko, Eugene Ilyushin [Paper]

  • MM-Soc: Benchmarking Multimodal Large Language Models in Social Media Platforms, Yiqiao Jin, Minje Choi, Gaurav Verma, Jindong Wang, Srijan Kumar [Paper]

  • Benchmarking Large Multimodal Models against Common Corruptions, Jiawei Zhang * 1 Tianyu Pang 2 Chao Du 2 Yi Ren † 3 Bo Li 1 4 Min Lin 2 [Paper] [Code]

Year 2023

  • Li, Yifan, et al. "Evaluating object hallucination in large vision-language models." arXiv preprint arXiv:2305.10355 (2023). [Paper] [Code]

  • Kamath, Amita, Jack Hessel, and Kai-Wei Chang. "What's" up" with vision-language models? Investigating their struggle with spatial reasoning." arXiv preprint arXiv:2310.19785 (2023). [Paper] [Code]