activity
20242026
collaborators

8 papers

cs.CV2026

h-Flow: Flexible Flow-based Image Editing via Doob's h-Transform

Zehui Guo, Zhen Wang, Junwei Shu +3

Editing images with pre-trained text-to-image flow models typically requires carefully balancing target alignment with the desired prompt and source consistency with the original i…

cs.CV2026

Bridging Rendering and Generative Modeling with Monte Carlo Transport Scheduling

Junwei Shu, Wenjie Liu, Hantang Liu +2

Monte Carlo rendering and modern generative models both transform uncertain states into structured images, yet they are usually studied as separate processes. We introduce Monte Ca…

cs.RO2026

AgentChemist: A Multi-Agent Experimental Robotic Platform Integrating Chemical Perception and Precise Control

Xiangyi Wei, Fei Wang, Haotian Zhang +6

Chemical laboratory automation has long been constrained by rigid workflows and poor adaptability to the long-tail distribution of experimental tasks. While most automated platform…

cs.CV2026

GT2-GS: Geometry-aware Texture Transfer for Gaussian Splatting

Wenjie Liu, Zhongliang Liu, Junwei Shu +2

Transferring 2D textures onto complex 3D scenes plays a vital role in enhancing the efficiency and controllability of 3D multimedia content creation. However, existing 3D style tra…

cs.RO2025

Audio-VLA: Adding Contact Audio Perception to Vision-Language-Action Model for Robotic Manipulation

Xiangyi Wei, Haotian Zhang, Xinyi Cao +4

The Vision-Language-Action models (VLA) have achieved significant advances in robotic manipulation recently. However, vision-only VLA models create fundamental limitations, particu…

cs.CV2025

LMP: Leveraging Motion Prior in Zero-Shot Video Generation with Diffusion Transformer

Changgu Chen, Xiaoyan Yang, Junwei Shu +2

In recent years, large-scale pre-trained diffusion transformer models have made significant progress in video generation. While current DiT models can produce high-definition, high…