2 papers
cs.CL2026
Shieldstral
Antonia Calvi, Avinash Sooriyarachchi, Giada Pistilli +274
We introduce Shieldstral, a 3B-parameter policy-adaptive multimodal safety classifier that matches or outperforms models nearly 7 its size on text safety benchmarks and set…
cs.CR2026
Mitigating Taint-Style Vulnerabilities in MCP Servers via Security-Aware Tool Descriptions
Yang Shi, Jiaheng Fu, Yihe Huang +3
Large language models (LLMs) are increasingly deployed as autonomous agents that interact with external tools and services via the Model Context Protocol (MCP), a standardized inte…