4 papers
U-Fold: Dynamic Intent-Aware Context Folding for User-Centric Agents
Jin Su, Runnan Fang, Yeqiu Li +5
Large language model (LLM)-based agents have been successfully deployed in many tool-augmented settings, but their scalability is fundamentally constrained by context length. Exist…
VARD: Efficient and Dense Fine-Tuning for Diffusion Models with Value-based RL
Fengyuan Dai, Zifeng Zhuang, Yufei Huang +4
Diffusion models have emerged as powerful generative tools across various domains, yet tailoring pre-trained models to exhibit specific desirable properties remains challenging. Wh…
Offline Trajectory Optimization for Offline Reinforcement Learning
Ziqi Zhao, Zhaochun Ren, Liu Yang +6
Offline reinforcement learning (RL) aims to learn policies without online explorations. To enlarge the training data, model-based offline RL learns a dynamics model which is utiliz…
Hypergraph Node Representation Learning with One-Stage Message Passing
Shilin Qu, Weiqing Wang, Yuan-Fang Li +2
Hypergraphs as an expressive and general structure have attracted considerable attention from various research domains. Most existing hypergraph node representation learning techni…