agent training 1container environments 1data generation 1large language models 1terminal task synthesis 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.AI2026
Meta-Task: Turning Terminal Task Synthesis into a Terminal Task for Scalable Agent Training
Zhihong Pan, Jiyuan He, Kai Zhang +5
The paper introduces Meta-Task, a framework that generates and verifies terminal tasks inside real container environments, enabling scalable training of terminal agents with high‑q…
cs.IR2025
From Entity Reliability to Clean Feedback: An Entity-Aware Denoising Framework Beyond Interaction-Level Signals
Ze Liu, Xianquan Wang, Shuochen Liu +5
Implicit feedback is central to modern recommender systems but is inherently noisy, often impairing model training and degrading user experience. At scale, such noise can mislead l…
cs.CL2025
Route to Reason: Adaptive Routing for LLM and Reasoning Strategy Selection
Zhihong Pan, Kai Zhang, Yuze Zhao +1
The inherent capabilities of a language model (LM) and the reasoning strategies it employs jointly determine its performance in reasoning tasks. While test-time scaling is regarded…