2 papers
cs.CL2026
Q-Interference: Memory-Efficient Phase-Aware Quantum-Inspired Attention
Emama Nahid, Tahmid Imtiaz Imu, Huayue Gu +3
GPT attention measures token compatibility through dot-product similarity. This mechanism is simple, effective, and memory-efficient. But it does not explicitly model whether stron…
cs.CR2026
COMIC: Reference-Aware Safety Gating for Multimodal Large Language Models
Md Abdullahil Oaphy, Anhao Xiang, Zongxing Xie +3
Multimodal large language models (MLLMs) are increasingly used to interact with screenshots, scanned documents, diagrams, and other visually grounded inputs. This shift introduces…