8 papers
h-Flow: Flexible Flow-based Image Editing via Doob's h-Transform
Zehui Guo, Zhen Wang, Junwei Shu +3
Editing images with pre-trained text-to-image flow models typically requires carefully balancing target alignment with the desired prompt and source consistency with the original i…
Bridging Rendering and Generative Modeling with Monte Carlo Transport Scheduling
Junwei Shu, Wenjie Liu, Hantang Liu +2
Monte Carlo rendering and modern generative models both transform uncertain states into structured images, yet they are usually studied as separate processes. We introduce Monte Ca…
AgentChemist: A Multi-Agent Experimental Robotic Platform Integrating Chemical Perception and Precise Control
Xiangyi Wei, Fei Wang, Haotian Zhang +6
Chemical laboratory automation has long been constrained by rigid workflows and poor adaptability to the long-tail distribution of experimental tasks. While most automated platform…
GT2-GS: Geometry-aware Texture Transfer for Gaussian Splatting
Wenjie Liu, Zhongliang Liu, Junwei Shu +2
Transferring 2D textures onto complex 3D scenes plays a vital role in enhancing the efficiency and controllability of 3D multimedia content creation. However, existing 3D style tra…
Audio-VLA: Adding Contact Audio Perception to Vision-Language-Action Model for Robotic Manipulation
Xiangyi Wei, Haotian Zhang, Xinyi Cao +4
The Vision-Language-Action models (VLA) have achieved significant advances in robotic manipulation recently. However, vision-only VLA models create fundamental limitations, particu…
LMP: Leveraging Motion Prior in Zero-Shot Video Generation with Diffusion Transformer
Changgu Chen, Xiaoyan Yang, Junwei Shu +2
In recent years, large-scale pre-trained diffusion transformer models have made significant progress in video generation. While current DiT models can produce high-definition, high…