collaborators

7 papers

cs.RO2026

HAM-VLN: Harnessing Hierarchical Agentic Memory for Zero-Shot Vision-and-Language Navigation

An Liu, Bingxi Liu, Hongyu Ding +6

Vision-and-language navigation (VLN) enables robots to follow instructions in previously unseen environments. Recently, a training-free paradigm has emerged: the robot queries a mu…

cs.LG2026

UNIFUSION: Adapting Autoregressive Language Models into Discrete Diffusion under a Unified Reverse-Rate Objective

Xiaoyi Jiang, Jingyuan Li, Yixuan Jiang +4

Existing methods mainly adapt pretrained autoregressive (AR) language models to masked diffusion, whereas we directly adapt them to uniform-noise diffusion, where every token remai…

cs.LG2026

Mean-to-Score Discrete Diffusion: Posterior-Mean Denoisers for Score Entropy

Jingyuan Li, Xiaoyi Jiang, Yixuan Jiang +4

Score Entropy Discrete Diffusion (SEDD) parameterizes discrete reverse processes with unconstrained positive score ratios. While positivity guarantees nonnegative reverse jump rate…

cs.LG2026

OLEDLM: A Unified Language Model for OLED Molecular Design

Fukang Wen, Yuchong Tang, Jingyuan Li +9

The development of organic light-emitting diode (OLED) materials faces the compounded challenges of an astronomically large chemical space, stringent quantum-chemical constraints,…

cs.RO2026

RealDexUMI: A Wearable Universal Manipulation Interface for Dexterous Robot Learning

Chaoyi Xu, Yixuan Jiang, Jiahui Huan +7

Learning dexterous manipulation requires demonstrations that preserve fine hand-object interactions while remaining executable at deployment. Existing pipelines either lose deploya…

q-bio.NC2025

Multiscale Causal Geometric Deep Learning for Modeling Brain Structure

Chengzhi Xia, Jianwei Chen, Yixuan Jiang +2

Multimodal MRI offers complementary multi-scale information to characterize the brain structure. However, it remains challenging to effectively integrate multimodal MRI while achie…