2 papers
cs.CV2026
ReactBench: A Cause-Driven Benchmark for Multimodal Hallucination via Systematic Evaluation
Shizhe Zhou, Bohan Jia, Kai Wu +4
While multimodal large language models (MLLMs) have achieved rapid progress in vision-language understanding, they remain prone to multimodal hallucinations, producing responses th…
cs.AI2024
The Llama 3 Herd of Models
Aaron Grattafiori, Abhimanyu Dubey, Abhinav Jauhri +556
Modern artificial intelligence (AI) systems are powered by foundation models. This paper presents a new set of foundation models, called Llama 3. It is a herd of language models th…