2 papers
cs.CL2026
Same Model, Different Weakness: How Language and Modality Reshape the Jailbreak Attack Surface in Frontier MLLMs
Casey Ford, Madison Van Doren, Sicheng Jin +1
The attack surface of a multimodal large language model (MLLM) is language-dependent in ways that reveal the mechanistic structure of alignment failures. We present the first syste…
cs.CL2026
"Be My Cheese?": Cultural Nuance Benchmarking for Machine Translation in Multilingual LLMs
Madison Van Doren, Casey Ford, Jennifer Barajas +2
We present a large-scale human evaluation benchmark for assessing cultural localisation in machine translation produced by state-of-the-art multilingual large language models (LLMs…