3 papers
cs.AI2026
AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction
Rui Zou, Yutao Zhu, Mengqi Wei +1
Large language models have demonstrated strong mathematical problem-solving capabilities, yet reliably verifying their candidate answers remains challenging. Existing representativ…
cs.AI2025
Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis
Rui Zou, Mengqi Wei, Yutao Zhu +3
Large Language Models (LLMs) excel in reasoning and generation across domains, but still struggle with identifying and diagnosing complex errors. This stems mainly from training ob…
cs.AI2024
Gradual Vigilance and Interval Communication: Enhancing Value Alignment in Multi-Agent Debates
Rui Zou, Mengqi Wei, Jintian Feng +3
In recent years, large language models have shown exceptional performance in fulfilling diverse human needs. However, their training data can introduce harmful content, underscorin…