19 citations · 19 across the 5 of their papers we have counts for
4 papers · 1 filter
CEDAR: Error-Bounded Residual Routing for Efficient Long-Context Attention
Siyu Li, Dong Wang, Jie Zhou +3
Post-hoc sparse attention accelerates long-context prefill by routing each query to a small set of token-level interactions. Hard selection, however, assigns zero probability to ev…
AutoMem: A Text-Gradient Recursive Self-Improvement Framework for Automated Memory Architectures Search
Lin Du, Jie Zhou, Yuxuan Cai +6
Long-term memory is increasingly central to LLM agents, yet memory design remains a highly coupled architecture problem: what to encode, how to store it, how to retrieve it, and ho…
OPV: Outcome-based Process Verifier for Efficient Long Chain-of-Thought Verification
Zijian Wu, Lingkai Kong, Wenwei Zhang +12
Large language models (LLMs) have achieved significant progress in solving complex reasoning tasks by Reinforcement Learning with Verifiable Rewards (RLVR). This advancement is als…
Investigating Public Fine-Tuning Datasets: A Complex Review of Current Practices from a Construction Perspective
Runyuan Ma, Wei Li, Fukai Shang
With the rapid development of the large model domain, research related to fine-tuning has concurrently seen significant advancement, given that fine-tuning is a constituent part of…