8 papers
Controllable Video Object Insertion via Multi-View Priors
Qi Xia, Xia Qi, Peishan Cong +4
Video object insertion places a user-specified object in an existing dynamic scene. Existing methods typically condition generation on text or a single reference image. Consequentl…
ReMoGen: Real-time Human Interaction-to-Reaction Generation via Modular Learning from Diverse Data
Yaoqin Ye, Yiteng Xu, Qin Sun +3
Human behaviors in real-world environments are inherently interactive, with an individual's motion shaped by surrounding agents and the scene. Such capabilities are essential for a…
In-Context Reinforcement Learning for Tool Use in Large Language Models
Yaoqi Ye, Yiran Zhao, Keyu Duan +4
While large language models (LLMs) exhibit strong reasoning abilities, their performance on complex tasks is often constrained by the limitations of their internal knowledge. A com…
ImageEdit-R1: Boosting Multi-Agent Image Editing via Reinforcement Learning
Yiran Zhao, Yaoqi Ye, Xiang Liu +2
With the rapid advancement of commercial multi-modal models, image editing has garnered significant attention due to its widespread applicability in daily life. Despite impressive…
LACONIC: Length-Aware Constrained Reinforcement Learning for LLM
Chang Liu, Yiran Zhao, Lawrence Liu +3
Reinforcement learning (RL) has enhanced the capabilities of large language models (LLMs) through reward-driven training. Nevertheless, this process can introduce excessively long…
Semiclassical analytical solutions of the eigenstate thermalization hypothesis in a quantum billiard
Yaoqi Ye, Chengkai Lin, Xiao Wang
We derive semiclassical analytical solutions for both the diagonal and off-diagonal functions in the eigenstate thermalization hypothesis (ETH) in a quarter-stadium quantum billiar…