3 papers
cs.LG2026
Train Less, Learn More: Adaptive Efficient Rollout Optimization for Group-Based Reinforcement Learning
Zhi Zhang, Zhen Han, Costas Mavromatis +9
Reinforcement learning (RL) plays a central role in large language model (LLM) post-training. Among existing approaches, Group Relative Policy Optimization (GRPO) is widely used, e…
cs.CL2025
Is Your LLM Outdated? A Deep Look at Temporal Generalization
Chenghao Zhu, Nuo Chen, Yufei Gao +3
The rapid advancement of Large Language Models (LLMs) has led to the development of benchmarks that consider temporal dynamics, however, there remains a gap in understanding how we…
cs.CL2024
ACER: Automatic Language Model Context Extension via Retrieval
Luyu Gao, Yunyi Zhang, Jamie Callan
Long-context modeling is one of the critical capabilities of language AI for digesting and reasoning over complex information pieces. In practice, long-context capabilities are typ…