From the 1 of 7 linked papers with an AI index.
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Agent Safety Is Action Alignment
Shawn Li, Yue Zhao
Large language models increasingly act as agents: they call tools, move money, delete records, and send messages on a user's behalf. To keep them safe, practitioners imported the c…
cs.AI2026
FORTIS: Benchmarking Over-Privilege in Agent Skills
Shawn Li, Chenxiao Yu, Han Wang +8
Large language model agents increasingly operate through an intermediate skill layer that mediates between user intent and concrete task execution. This layer is widely treated as…
cs.AI2026
Geometry over Density: Few-Shot Cross-Domain OOD Detection
Shawn Li, You Qin, Jiate Li +4
Out-of-distribution (OOD) detection identifies test samples that fall outside a model's training distribution, a capability critical for safe deployment in high-stakes applications…