2 papers
cs.CL2025
Judge Before Answer: Can MLLM Discern the False Premise in Question?
Jidong Li, Lingyong Fang, Haodong Zhao +2
Multimodal large language models (MLLMs) have witnessed astonishing advancements in recent years. Despite these successes, MLLMs remain vulnerable to flase premise problems. Howeve…
cs.AI2025
NCV: A Node-Wise Consistency Verification Approach for Low-Cost Structured Error Localization in LLM Reasoning
Yulong Zhang, Li Wang, Wei Du +7
Verifying multi-step reasoning in large language models is difficult due to imprecise error localization and high token costs. Existing methods either assess entire reasoning chain…