2 papers
cs.CL2026
SSA: Improving Performance With a Better Scoring Function
Omar Naim, Swarnadeep Bhar, Jérôme Bolte +1
While transformer models exhibit strong in-context learning (ICL) abilities, they often fail to generalize under simple distribution shifts. We analyze these failures and identify…
cs.CL2026
COCORELI: Enforcing Execution Preconditions for Reliable Collaborative Instruction Following
Swarnadeep Bhar, Omar Naim, Eleni Metheniti +4
Autonomous agents executing human instructions must operate reliably even when instructions are incomplete. While recent approaches improve detection of missing information, detect…