3 papers
cs.AI2026
When Words Are Safe But Actions Kill: Probing Physical Jailbreak Beyond Textual Jailbreak in Hidden-State Risk Space
Weimeng Wang, Ziqiang Wang, Zihang Zhan +3
Large language models (LLMs) increasingly serve as high-level planners for embodied agents, where linguistically benign instructions can become unsafe once grounded in the physical…
cs.CL2026
Evaluating and Calibrating LLM Confidence on Questions with Multiple Correct Answers
Yuhan Wang, Shiyu Ni, Zhikai Ding +3
Confidence calibration is essential for making large language models (LLMs) reliable, yet existing training-free methods have been primarily studied under single-answer question an…
cs.HC2026
Earinter: A Closed-Loop System for Eating Pace Regulation with Just-in-Time Intervention Using Commodity Earbuds
Jun Fang, Ka I Chan, Xiyuxing Zhang +7
Rapid eating is common yet difficult to regulate in situ, partly because people seldom notice pace changes and sustained self-monitoring is effortful. We present Earinter, a commod…