2 papers
cs.CV2025
ViBe: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models
Vipula Rawte, Sarthak Jain, Aarush Sinha +9
Recent advances in Large Multimodal Models (LMMs) have expanded their capabilities to video understanding, with Text-to-Video (T2V) models excelling in generating videos from textu…
cs.CL2024
GRS-QA -- Graph Reasoning-Structured Question Answering Dataset
Anish Pahilajani, Devasha Trivedi, Jincen Shuai +7
Large Language Models (LLMs) have excelled in multi-hop question-answering (M-QA) due to their advanced reasoning abilities. However, the impact of the inherent reasoning structure…