2 papers
cs.CV2026
CinematicVQA: Benchmarking Film-Grammar Reasoning in Large Vision-Language Models
Shuo Xing, Pooja Verlani, Balu Adsumilli +1
Cinematography, the craft of visual storytelling through framing, lighting, and camera operation, fundamentally shapes how audiences perceive and emotionally engage with video cont…
cs.CV2026
SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation
Jiongze Yu, Xiangbo Gao, Pooja Verlani +4
Video Super-Resolution (VSR) aims to restore high-quality video frames from low-resolution (LR) estimates, yet most existing VSR approaches behave like black boxes at inference tim…