5 papers · 1 filter
Delving Deep into Engagement Prediction of Short Videos
Dasong Li, Wenjie Li, Baili Lu +4
Understanding and modeling the popularity of User Generated Content (UGC) short videos on social media platforms presents a critical challenge with broad implications for content c…
T2M-X: Learning Expressive Text-to-Motion Generation from Partially Annotated Data
Mingdian Liu, Yilin Liu, Gurunandan Krishnan +2
The generation of humanoid animation from text prompts can profoundly impact animation production and AR/VR experiences. However, existing methods only generate body motion data, e…
DSL-FIQA: Assessing Facial Image Quality via Dual-Set Degradation Learning and Landmark-Guided Transformer
Wei-Ting Chen, Gurunandan Krishnan, Qiang Gao +3
Generic Face Image Quality Assessment (GFIQA) evaluates the perceptual quality of facial images, which is crucial in improving image restoration algorithms and selecting high-quali…
Personalized Restoration via Dual-Pivot Tuning
Pradyumna Chari, Sizhuo Ma, Daniil Ostashev +4
Generative diffusion models can serve as a prior which ensures that solutions of image restoration systems adhere to the manifold of natural images. However, for restoring facial i…
Velocity Disambiguation for Video Frame Interpolation
Zhihang Zhong, Yiming Zhang, Wei Wang +5
Existing video frame interpolation (VFI) methods blindly predict where each object is at a specific timestep t ("time indexing"), which struggles to predict precise object movement…