2 papers
cs.LG2025
Model Extrapolation Expedites Alignment
Chujie Zheng, Ziqi Wang, Heng Ji +2
Given the high computational cost of preference alignment training of large language models (LLMs), exploring efficient methods to reduce the training overhead remains an important…
cs.CL2024
AMOR: A Recipe for Building Adaptable Modular Knowledge Agents Through Process Feedback
Jian Guan, Wei Wu, Zujie Wen +3
The notable success of large language models (LLMs) has sparked an upsurge in building language agents to complete various complex tasks. We present AMOR, an agent framework based…