3 papers
cs.LG2025
CrystalDiT: A Diffusion Transformer for Crystal Generation
Xiaohan Yi, Guikun Xu, Xi Xiao +4
We present CrystalDiT, a diffusion transformer for crystal structure generation that achieves state-of-the-art performance by challenging the trend of architectural complexity. Ins…
cs.CV2025
PIG-Nav: Key Insights for Pretrained Image Goal Navigation Models
Jiansong Wan, Chengming Zhou, Jinkua Liu +14
Recent studies have explored pretrained (foundation) models for vision-based robotic navigation, aiming to achieve generalizable navigation and positive transfer across diverse env…
cs.RO2024
UniGraspTransformer: Simplified Policy Distillation for Scalable Dexterous Robotic Grasping
Wenbo Wang, Fangyun Wei, Lei Zhou +9
We introduce UniGraspTransformer, a universal Transformer-based network for dexterous robotic grasping that simplifies training while enhancing scalability and performance. Unlike…