5 papers
ChainClaw: A Layered Agent Framework for Reliable On-Chain Execution
Jiacheng Wei, Zhaoxin Fan, Xin Wen +5
General-purpose large language model agents have achieved strong performance on tool-augmented tasks, yet they rely on assumptions break down in blockchain environments. On-chain e…
PACT: Learning Diverse Diagnostic Strategies via Privileged Synthesis and Branch Consensus
Gen Li, Yuanze Hu, Zhichao Yang +10
Clinical diagnosis requires flexible use of multiple reasoning paradigms under incomplete patient information. Existing LLM-based medical agents show strong medical reasoning abili…
State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading
Yuanze Hu, Gen Li, Yuqin Lan +5
Multimodal large language models (MLLMs) have achieved impressive progress on general multimodal tasks, yet they remain brittle on dial-based measurement reading. In this paper, we…
Lyapunov Probes for Hallucination Detection in Large Foundation Models
Bozhi Luan, Gen Li, Yalan Qin +6
We address hallucination detection in Large Language Models (LLMs) and Multimodal Large Language Models (MLLMs) by framing the problem through the lens of dynamical systems stabili…
EMG-UP: Unsupervised Personalization in Cross-User EMG Gesture Recognition
Nana Wang, Suli Wang, Gen Li +1
Cross-user electromyography (EMG)-based gesture recognition represents a fundamental challenge in achieving scalable and personalized human-machine interaction within real-world ap…