1 citations · 1 across the 12 of their papers we have counts for
13 papers · 1 filter
LumaGuide: Distribution Shaping for Training-Free HDR Generation in Diffusion Models
Bowen Chen, Shreshth Saini, Balu Adsumilli +1
Pretrained diffusion models generate realistic images but are constrained by the statistical biases of their training data, limiting their ability to produce high dynamic range (HD…
RUTA: Principled Visual Token Allocation via Rate-Utility Optimization
Jian Zou, Xiaoyu Xu, Zhihua Wang +3
High-resolution images and long videos provide vision-language models with rich context for multimodal reasoning and fine-grained perception, but the resulting long visual token se…
Subjective Portrait Region Cropping in Landscape Videos with Temporal Annotation Smoothing
Cheng-Han Lee, Maniratnam Mandal, Neil Birkbeck +3
With the rise of mobile video consumption on diverse handheld display resolutions and orientation modes, altering videos to aspect ratios poses challenges. Static cropping and bord…
LumaFlux: Lifting 8-Bit Worlds to HDR Reality with Physically-Guided Diffusion Transformers
Shreshth Saini, Hakan Gedik, Neil Birkbeck +3
The rapid adoption of HDR-capable devices has created a pressing need to convert the 8-bit Standard Dynamic Range (SDR) content into perceptually and physically accurate 10-bit Hig…
SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation
Jiongze Yu, Xiangbo Gao, Pooja Verlani +4
Video Super-Resolution (VSR) aims to restore high-quality video frames from low-resolution (LR) estimates, yet most existing VSR approaches behave like black boxes at inference tim…
MDS-VQA: Model-Informed Data Selection for Video Quality Assessment
Jian Zou, Xiaoyu Xu, Zhihua Wang +3
Learning-based video quality assessment (VQA) has advanced rapidly, yet progress is increasingly constrained by a disconnect between model design and dataset curation. Model-centri…