3 papers
cs.RO2025
DM1: MeanFlow with Dispersive Regularization for 1-Step Robotic Manipulation
Guowei Zou, Haitao Wang, Hejun Wu +3
The ability to learn multi-modal action distributions is indispensable for robotic manipulation policies to perform precise and robust control. Flow-based generative models have re…
cs.CV2025
Preserve and Sculpt: Manifold-Aligned Fine-tuning of Vision-Language Models for Few-Shot Learning
Dexia Chen, Qianjie Zhu, Weibing Li +3
Pretrained vision-language models (VLMs), such as CLIP, have shown remarkable potential in few-shot image classification and led to numerous effective transfer learning strategies.…
cs.CV2025
Cross-Domain Few-Shot Learning via Multi-View Collaborative Optimization with Vision-Language Models
Dexia Chen, Wentao Zhang, Qianjie Zhu +4
Vision-language models (VLMs) pre-trained on natural image and language data, such as CLIP, have exhibited significant potential in few-shot image recognition tasks, leading to dev…