2 papers
cs.CV2026
INFACT: A Diagnostic Benchmark for Induced Faithfulness and Factuality Hallucinations in Video-LLMs
Junqi Yang, Yuecong Min, Jie Zhang +2
Despite rapid progress, Video Large Language Models (Video-LLMs) remain unreliable due to hallucinations, which are outputs that contradict either video evidence (faithfulness) or…
cs.LG2025
Beyond Frequency: The Role of Redundancy in Large Language Model Memorization
Jie Zhang, Qinghua Zhao, Chi-ho Lin +2
Memorization in large language models poses critical risks for privacy and fairness as these systems scale to billions of parameters. While previous studies established correlation…