2 papers
cs.IR2026
MISO: Model-Internal-State-Guided Optimization for Ranking Models
Yongzhe Zhang, Xiaoyu Deng, Yifan He +28
Ranking models are repeatedly refined within established model families, yet the choice of which component to scale, replace, or retire is often guided by expensive trial-and-error…
cs.AI2026
Long-Horizon Agent Trajectory Attribution: A Unified Benchmark and Fine-Grained Annotation Framework
Jing Chen, Yang Sun, Li Zhang +2
Large language model (LLM) agents increasingly operate through long-horizon trajectories involving user instructions, tool use, external observations, and memory. Existing benchmar…