4 papers
LiveMem: Maintaining Memory State Continuity in Long-Running LLM Inference
Zhichen Liu, Ruihan Sun, Hengjie Yang +4
Long-running assistants and agents consume interaction streams that eventually outgrow the context. Existing context retention, summarization, and retrieval preserve access to sele…
X4Val: Learning Neural Surrogates for Variance-Reduced Policy Evaluation
Rachel Luo, Michael Watson, Apoorva Sharma +6
Rigorous evaluation of learning-based robotic systems is an essential prerequisite for deployment. However, real-world test data is expensive to gather; moreover, in a typical iter…
Flexible Locomotion Learning with Diffusion Model Predictive Control
Runhan Huang, Haldun Balim, Heng Yang +1
Legged locomotion demands controllers that are both robust and adaptable, while remaining compatible with task and safety considerations. However, model-free reinforcement learning…
SPIE: Semantic and Structural Post-Training of Image Editing Diffusion Models with AI feedback
Elior Benarous, Yilun Du, Heng Yang
This paper presents SPIE: a novel approach for semantic and structural post-training of instruction-based image editing diffusion models, addressing key challenges in alignment wit…