action recognition 1disentanglement 1knowledge-guided learning 1large language models 1multi-label classification 1video understanding 1
From the 1 of 3 linked papers with an AI index.
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Knowledge-guided Disentanglement with Atomic Actions for Action Recognition
Tianci Wu, Siqi Cao, Guangming Zhu +6
The paper introduces a framework that uses large language models to break down action labels into atomic actions and injects this semantic knowledge into video features to improve…
cs.CV2026
Look Clearly Before Answering: Mitigating Hallucinations in LVLMs via Saliency-Driven Perceptual Realignment
Pengxu Chen, Yao Zhu, Guangming Zhu +4
Large vision-language models (LVLMs) have demonstrated remarkable capabilities in multimodal understanding. However, they remain prone to hallucinations, generating responses that…