3 papers
cs.AI2026
Buried in Textual Debt: Context Pruning with Visual Evidence Preservation for MLLM Agents
Yuchen Huang, Sijia Li, Jun Zhang +1
Multimodal Large Language Models (MLLMs) are increasingly deployed as multi-step agents, where explicit reasoning supports task decomposition and tool coordination but also accumul…
cs.AI2026
CDR-Bench: Evaluating Faithful Execution of Compositional, Order-Sensitive Data Refinement Recipes
Yuchen Huang, Xiang Li, Zhenqing Ling +5
Data refinement involves executing multi-step recipes over evolving text states, where both composition and execution order of processing operators determine the outcome. While exi…
cs.LG2025
Environment Scaling for Interactive Agentic Experience Collection: A Survey
Yuchen Huang, Sijia Li, Minghao Liu +5
LLM-based agents can autonomously accomplish complex tasks across various domains. However, to further cultivate capabilities such as adaptive behavior and long-term decision-makin…