collaborators

23 papers

cs.AI2026

Quantifying Consistency in LLM Logical Reasoning via Structural Uncertainty

Baishali Chaudhury, Mengdie Flora Wang, Hyunji Hayley Park +3

Large language models can arrive at the same answer through reasoning paths that are unstable, contradictory, or difficult to rank consistently -- a failure mode especially prevale…

cs.AI2026

Knowing When to Ask: Self-Gated Clarification for Hierarchical Language Agents

Aijing Gao, Yiming Kang, Mengdie Flora Wang +1

In hierarchical reasoning, failures often originate at intermediate decision points where the agent commits to a wrong branch without recognizing that it lacks critical information…

physics.chem-ph2026

LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning

Xinwu Ye, Yicheng Mao, Yuxuan Liao +16

Current chemical large language models (LLMs) predominantly rely on explicit Chain-of-Thought (CoT) to solve complex reasoning problems. However, forcing nonverbal tacit chemical l…

cs.CL2026

Learning Agent Routing From Early Experience

Yimin Wang, Jiahao Qiu, Xuan Qi +6

LLM agents achieve strong performance on complex reasoning tasks but incur high latency and compute cost. In practice, many queries fall within the capability boundary of cutting-e…

cs.AI2026

Position: Agent Should Invoke External Tools ONLY When Epistemically Necessary

Hongru Wang, Cheng Qian, Manling Li +6

As large language models evolve into tool-augmented agents, a central question remains unresolved: when is external tool use actually justified? Existing agent frameworks typically…

cs.CL2026

From Word to World: Can Large Language Models be Implicit Text-based World Models?

Yixia Li, Hongru Wang, Jiahao Qiu +7

Agentic reinforcement learning increasingly relies on experience-driven scaling, yet real-world environments remain non-adaptive, limited in coverage, and difficult to scale. World…