papers

Publications (10)

cs.MM2019

MMED: A Multi-domain and Multi-modality Event Dataset

Zhenguo Yang, Zehang Lin, Min Cheng +2

In this work, we construct and release a multi-domain and multi-modality event dataset (MMED), containing 25,165 textual news articles collected from hundreds of news media sites (…

physics.comp-ph2004

Improved neighbor list algorithm in molecular simulations using cell decomposition and data sorting method

Zhenhua Yao, Jian-Sheng Wang, Gui-Rong Liu +1

An improved neighbor list algorithm is proposed to reduce unnecessary interatomic distance calculations in molecular simulations. It combines the advantages of Verlet table and cel…

cs.RO2026

Learning Surgical Robotic Manipulation with 3D Spatial Priors

Yu Sheng, Lidian Wang, Xiaomeng Chu +6

Achieving 3D spatial awareness is crucial for surgical robotic manipulation, where precise and delicate operations are required. Existing methods either explicitly reconstruct the…

cs.NI2006

Grooming of Dynamic Traffic in WDM Star and Tree Networks Using Genetic Algorithm

Kun-hong Liu, Yong Xu, De-shuang Huang +1

The advances in WDM technology lead to the great interest in traffic grooming problems. As traffic often changes from time to time, the problem of grooming dynamic traffic is of gr…

cs.LG2023

Natural Actor-Critic for Robust Reinforcement Learning with Function Approximation

Ruida Zhou, Tao Liu, Min Cheng +3

We study robust reinforcement learning (RL) with the goal of determining a well-performing policy that is robust against model mismatch between the training simulator and the testi…

cs.RO2026

Open-H-Embodiment: A Large-Scale Dataset for Enabling Foundation Models in Medical Robotics

Open-H-Embodiment Consortium, :, Nigel Nelson +213

Autonomous medical robots hold promise to improve patient outcomes, reduce provider workload, democratize access to care, and enable superhuman precision. However, autonomous medic…

cs.CV2025

DCL-SE: Dynamic Curriculum Learning for Spatiotemporal Encoding of Brain Imaging

Meihua Zhou, Xinyu Tong, Jiarui Zhao +4

High-dimensional neuroimaging analyses for clinical diagnosis are often constrained by compromises in spatiotemporal fidelity and by the limited adaptability of large-scale, genera…

cs.LG2024

Provable Policy Gradient Methods for Average-Reward Markov Potential Games

Min Cheng, Ruida Zhou, P. R. Kumar +1

We study Markov potential games under the infinite horizon average reward criterion. Most previous studies have been for discounted rewards. We prove that both algorithms based on…

cs.AI2026

Diffusion Blend: Inference-Time Multi-Preference Alignment for Diffusion Models

Min Cheng, Fatemeh Doudi, Dileep Kalathil +2

Reinforcement learning (RL) algorithms have been used recently to align diffusion models with downstream objectives such as aesthetic quality and text-image consistency by fine-tun…

cs.AI2026

Long-Horizon Plan Execution in Large Tool Spaces through Entropy-Guided Branching

Rongzhe Wei, Ge Shi, Min Cheng +5

Large Language Models (LLMs) have significantly advanced tool-augmented agents, enabling autonomous reasoning via API interactions. However, executing multi-step tasks within massi…