4 papers
Budgeted Act-or-Defer Multi-Agent LLM Deliberation with Local Reliability Bounds
Mengdie Flora Wang, Haochen Xie, Guanghui Wang +2
Multi-agent deliberation among LLMs can improve reasoning, but deployment requires deciding when the current answer is reliable enough to act on and when it should be escalated to…
Quantifying Consistency in LLM Logical Reasoning via Structural Uncertainty
Baishali Chaudhury, Mengdie Flora Wang, Hyunji Hayley Park +3
Large language models can arrive at the same answer through reasoning paths that are unstable, contradictory, or difficult to rank consistently -- a failure mode especially prevale…
Knowing When to Ask: Self-Gated Clarification for Hierarchical Language Agents
Aijing Gao, Yiming Kang, Mengdie Flora Wang +1
In hierarchical reasoning, failures often originate at intermediate decision points where the agent commits to a wrong branch without recognizing that it lacks critical information…
From Debate to Decision: Conformal Social Choice for Safe Multi-Agent Deliberation
Mengdie Flora Wang, Haochen Xie, Guanghui Wang +9
Multi-agent debate improves LLM reasoning, yet agreement among agents is not evidence of correctness. When agents converge on a wrong answer through social reinforcement, consensus…