2 citations · 3 across the 4 of their papers we have counts for
8 papers · 1 filter
When Distillation Breaks Motion Control: Restoring Generative Trajectories for Fast Video Generators
Jintao Rong, Xin Xie, Xinyi Yu +4
Training-free motion customization imposes motion patterns from reference videos onto video generators through test-time computation. Most existing methods target full diffusion mo…
Boosting Box-supervised Instance Segmentation with Pseudo Depth
Xinyi Yu, Ling Yan, Pengtao Jiang +4
The realm of Weakly Supervised Instance Segmentation (WSIS) under box supervision has garnered substantial attention, showcasing remarkable advancements in recent years. However, t…
Improving Neural Indoor Surface Reconstruction with Mask-Guided Adaptive Consistency Constraints
Xinyi Yu, Liqin Lu, Jintao Rong +2
3D scene reconstruction from 2D images has been a long-standing task. Instead of estimating per-frame depth maps and fusing them in 3D, recent research leverages the neural implici…
ShiftNAS: Improving One-shot NAS via Probability Shift
Mingyang Zhang, Xinyi Yu, Haodong Zhao +1
One-shot Neural architecture search (One-shot NAS) has been proposed as a time-efficient approach to obtain optimal subnet architectures and weights under different complexity case…
Retrieval-Enhanced Visual Prompt Learning for Few-shot Classification
Jintao Rong, Hao Chen, Linlin Ou +3
The Contrastive Language-Image Pretraining (CLIP) model has been widely used in various downstream vision tasks. The few-shot learning paradigm has been widely adopted to augment i…
CrossFusion: Interleaving Cross-modal Complementation for Noise-resistant 3D Object Detection
Yang Yang, Weijie Ma, Hao Chen +2
The combination of LiDAR and camera modalities is proven to be necessary and typical for 3D object detection according to recent studies. Existing fusion strategies tend to overly…