9 papers
A Measure-Theoretic Analysis of Reasoning: Structural Generalization and Approximation Limits
Yuyang Zhang, Yifu Zhang, Xuehai Zhou +1
While empirical scaling laws for LLM reasoning are well-documented, the theoretical mechanisms governing out-of-distribution (OOD) generalization remain elusive. We formalize reaso…
LiFT: Lifted Inter-slice Feature Trajectories for 3D Image Generation from 2D Generators
Xinhe Zhang, Yuyang Zhang, Pengfei Jin +3
High-resolution 3D medical image generation remains challenging because fully volumetric models are computationally expensive, while efficient 2D slice generators often fail to pre…
Provably Efficient Sensor Allocation for Unknown High-dimensional Systems with Limited Sensing
Yuyang Zhang, Derya Cansever, Na Li
This paper focuses on learning efficient sensor allocations that ensure observability of unknown high-dimensional linear systems using only a small number of sensors. Existing meth…
Decentralized Diffusion Policy Learning for Enhanced Exploration in Cooperative Multi-agent Reinforcement Learning
Yuyang Zhang, Haldun Balim, Na Li
Cooperative multi-agent reinforcement learning (MARL) involves complex agent interactions and requires effective exploration strategies. A prominent class of MARL algorithms, decen…
Human-AI Co-Embodied Intelligence for Scientific Experimentation and Manufacturing
Xinyi Lin, Yuyang Zhang, Yuanhang Gan +10
Scientific experimentation and manufacturing rely on prolonged protocol development and complex, multi-step implementation, which require continuous human expertise for precise exe…
Max-Entropy Reinforcement Learning with Flow Matching and A Case Study on LQR
Yuyang Zhang, Yang Hu, Bo Dai +1
Soft actor-critic (SAC) is a popular algorithm for max-entropy reinforcement learning. In practice, the energy-based policies in SAC are often approximated using simple policy clas…