16 citations · 30 across the 4 of their papers we have counts for
11 papers · 1 filter
Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling
Gongye Liu, Bo Yang, Yida Zhi +8
Preference optimization for diffusion and flow-matching models relies on reward functions that are both discriminatively robust and computationally efficient. Vision-Language Model…
BiVM: Accurate Binarized Neural Network for Efficient Video Matting
Haotong Qin, Xianglong Liu, Xudong Ma +4
Deep neural networks for real-time video matting suffer significant computational limitations on edge devices, hindering their adoption in widespread applications such as online co…
Matching Anything by Segmenting Anything
Siyuan Li, Lei Ke, Martin Danelljan +4
The robust association of the same objects across video frames in complex scenes is crucial for many applications, especially Multiple Object Tracking (MOT). Current methods predom…
DreamScene4D: Dynamic Multi-Object Scene Generation from Monocular Videos
Wen-Hsuan Chu, Lei Ke, Katerina Fragkiadaki
View-predictive generative models provide strong priors for lifting object-centric images and videos into 3D and 4D through rendering and score distillation objectives. A question…
Occlusion-Aware Video Object Inpainting
Lei Ke, Yu-Wing Tai, Chi-Keung Tang
Conventional video inpainting is neither object-oriented nor occlusion-aware, making it liable to obvious artifacts when large occluded object regions are inpainted. This paper pre…
Deep Occlusion-Aware Instance Segmentation with Overlapping BiLayers
Lei Ke, Yu-Wing Tai, Chi-Keung Tang
Segmenting highly-overlapping objects is challenging, because typically no distinction is made between real object contours and occlusion boundaries. Unlike previous two-stage inst…