4 papers · 1 filter
When Distillation Breaks Motion Control: Restoring Generative Trajectories for Fast Video Generators
Jintao Rong, Xin Xie, Xinyi Yu +4
Training-free motion customization imposes motion patterns from reference videos onto video generators through test-time computation. Most existing methods target full diffusion mo…
Retrieval-Enhanced Visual Prompt Learning for Few-shot Classification
Jintao Rong, Hao Chen, Linlin Ou +3
The Contrastive Language-Image Pretraining (CLIP) model has been widely used in various downstream vision tasks. The few-shot learning paradigm has been widely adopted to augment i…
InstantStyleGaussian: Efficient Art Style Transfer with 3D Gaussian Splatting
Xin-Yi Yu, Jun-Xin Yu, Li-Bo Zhou +2
We present InstantStyleGaussian, an innovative 3D style transfer method based on the 3D Gaussian Splatting (3DGS) scene representation. By inputting a target-style image, it quickl…
Boosting Box-supervised Instance Segmentation with Pseudo Depth
Xinyi Yu, Ling Yan, Pengtao Jiang +4
The realm of Weakly Supervised Instance Segmentation (WSIS) under box supervision has garnered substantial attention, showcasing remarkable advancements in recent years. However, t…