3 papers
cs.RO2026
WISE: World-model-guided Imagination Scheduling for Efficient Post-training of Vision-Language-Action Models
Chenhao Zhang, Hanyu Zhao, Hang Cheng +2
Post-training VLA policies typically rely on supervised fine-tuning with costly expert demonstrations or reinforcement learning with expensive and potentially unstable real-world e…
cs.CV2026
Diff-SBSR: Learning Multimodal Feature-Enhanced Diffusion Models for Zero-Shot Sketch-Based 3D Shape Retrieval
Hang Cheng, Fanhe Dong, Long Zeng
This paper presents the first exploration of text-to-image diffusion models for zero-shot sketch-based 3D shape retrieval (ZS-SBSR). Existing sketch-based 3D shape retrieval method…
cs.CV2026
Multi-View Hierarchical Graph Neural Network for Sketch-Based 3D Shape Retrieval
Hang Cheng, Muyan He, Mingyu Fan +3
Sketch-based 3D shape retrieval (SBSR) aims to retrieve 3D shapes that are consistent with the category of the input hand-drawn sketch. The core challenge of this task lies in two…