16 papers
Pro-Pose: Unpaired Full-Body Portrait Synthesis via Canonical UV Maps
Sandeep Mishra, Yasamin Jafarian, Andreas Lugmayr +5
Photographs of people taken by professional photographers typically present the person in beautiful lighting, with an interesting pose, and flattering quality. This is unlike commo…
TABES: Trajectory-Aware Backward-on-Entropy Steering for Masked Diffusion Models
Shreshth Saini, Avinab Saha, Balu Adsumilli +3
Masked Diffusion Models (MDMs) have emerged as a promising non-autoregressive paradigm for generative tasks, offering parallel decoding and bidirectional context utilization. Howev…
Constructing Per-Shot Bitrate Ladders using Visual Information Fidelity
Krishna Srikar Durbha, Alan C. Bovik
Video service providers need their delivery systems to be able to adapt to network conditions, user preferences, display settings, and other factors. HTTP Adaptive Streaming (HAS)…
VIDMP3: Video Editing by Representing Motion with Pose and Position Priors
Sandeep Mishra, Oindrila Saha, Alan C. Bovik
Motion-preserved video editing is crucial for creators, particularly in scenarios that demand flexibility in both the structure and semantics of swapped objects. Despite its potent…
CHUG: Crowdsourced User-Generated HDR Video Quality Dataset
Shreshth Saini, Alan C. Bovik, Neil Birkbeck +2
High Dynamic Range (HDR) videos enhance visual experiences with superior brightness, contrast, and color depth. The surge of User-Generated Content (UGC) on platforms like YouTube…
PIT-QMM: A Large Multimodal Model For No-Reference Point Cloud Quality Assessment
Shashank Gupta, Gregoire Phillips, Alan C. Bovik
Large Multimodal Models (LMMs) have recently enabled considerable advances in the realm of image and video quality assessment, but this progress has yet to be fully explored in the…