activity
20242026
collaborators

9 papers

cs.CV2026

HPSv3++: Scaling Reward Models Across the Full Spectrum of Diffusion Model Capabilities

Yijun Liu, Jie Huang, Zeyue Xue +5

Reward models guide text-to-image (T2I) systems toward outputs aligned with human preferences. However, typical reward models such as HPSv3 are trained on pre-annotated data from e…

cs.CV2026

MesonGS++: Post-training Compression of 3D Gaussian Splatting with Hyperparameter Searching

Shuzhao Xie, Junchen Ge, Weixiang Zhang +10

3D Gaussian Splatting (3DGS) achieves high-quality novel view synthesis with real-time rendering, but its storage cost remains prohibitive for practical deployment. Existing post-t…

cs.CV2025

SizeGS: Size-aware Compression of 3D Gaussian Splatting via Mixed Integer Programming

Shuzhao Xie, Jiahang Liu, Weixiang Zhang +7

Recent advances in 3D Gaussian Splatting (3DGS) have greatly improved 3D reconstruction. However, its substantial data size poses a significant challenge for transmission and stora…

cs.RO2025

Spatial Policy: Guiding Visuomotor Robotic Manipulation with Spatial-Aware Modeling and Reasoning

Yijun Liu, Yuwei Liu, Yuan Meng +8

Vision-centric hierarchical embodied models have demonstrated strong potential. However, existing methods lack spatial awareness capabilities, limiting their effectiveness in bridg…

cs.RO2025

VGGT-DP: Generalizable Robot Control via Vision Foundation Models

Shijia Ge, Yinxin Zhang, Shuzhao Xie +3

Visual imitation learning frameworks allow robots to learn manipulation skills from expert demonstrations. While existing approaches mainly focus on policy design, they often negle…

cs.CV2025

Expansive Supervision for Neural Radiance Field

Weixiang Zhang, Shuzhao Xie, Shijia Ge +3

Neural Radiance Field (NeRF) has achieved remarkable success in creating immersive media representations through its exceptional reconstruction capabilities. However, the computati…