Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
A Survey: Spatiotemporal Consistency in Video Generation
Zhiyu Yin, Kehai Chen, Xuefeng Bai +7
Video generation aims to produce temporally coherent sequences of visual frames, representing a pivotal advancement in Artificial Intelligence Generated Content (AIGC). Compared to…
cs.CV2026
Beyond the Vision Encoder: Identifying and Mitigating Spatial Bias in Large Vision-Language Models
Yingjie Zhu, Xuefeng Bai, Kehai Chen +6
Large Vision-Language Models (LVLMs) have achieved remarkable success across a wide range of multimodal tasks, yet their robustness to spatial variations remains insufficiently und…