1 citations · 1 across the 9 of their papers we have counts for
10 papers
VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection
Haowen Sun, Shaolong Zhang, Mingyang Li +6
Open-vocabulary 3D affordance detection requires localizing interaction regions on point clouds given novel affordance descriptions. Recent methods extend multimodal large language…
GeoBlock: Inferring Block Granularity from Dependency Geometry in Diffusion Language Models
Lipeng Wan, Junjie Ma, Jianhui Gu +3
Block diffusion enables efficient parallel refinement in diffusion language models, but its decoding behavior depends critically on block size. Existing block-sizing strategies rel…
Progressive Refinement Regulation for Accelerating Diffusion Language Model Decoding
Lipeng Wan, Jianhui Gu, Junjie Ma +4
Diffusion language models generate text through iterative denoising under a uniform refinement rule applied to all tokens. However, tokens stabilize at different rates in practice,…
Bridging Simulation and Reality: Cross-Domain Transfer with Semantic 2D Gaussian Splatting
Jian Tang, Pu Pang, Haowen Sun +4
Cross-domain transfer in robotic manipulation remains a longstanding challenge due to the significant domain gap between simulated and real-world environments. Existing methods suc…
PRISM: Projection-based Reward Integration for Scene-Aware Real-to-Sim-to-Real Transfer with Few Demonstrations
Haowen Sun, Han Wang, Chengzhong Ma +4
Learning from few demonstrations to develop policies robust to variations in robot initial positions and object poses is a problem of significant practical interest in robotics. Co…
Playing Non-Embedded Card-Based Games with Reinforcement Learning
Tianyang Wu, Lipeng Wan, Yuhang Wang +2
Significant progress has been made in AI for games, including board games, MOBA, and RTS games. However, complex agents are typically developed in an embedded manner, directly acce…