Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Exact Is Easier: Credit Assignment for Cooperative LLM Agents
Yanjun Chen, Yirong Sun, Hanlin Wang +5
Removing an agent from a cooperative team to measure its contribution seems natural, yet in multi-agent LLM systems this evaluation distorts the result it claims to measure. This f…
cs.LG2025
STeCa: Step-level Trajectory Calibration for LLM Agent Learning
Hanlin Wang, Jian Wang, Chak Tou Leong +1
Large language model (LLM)-based agents have shown promise in tackling complex tasks by interacting dynamically with the environment. Existing work primarily focuses on behavior cl…