2 papers
cs.CL2026
Can LLMs Reliably Self-Report Adversarial Prefills, and How?
Quang Minh Nguyen, Uzair Ahmed, Taegyoon Kim
Prior work shows that large language models (LLMs) exhibit introspective capability on benign tasks. We extend the question to safety contexts and examine how reliably a model can…
cs.AI2025
SOLAR: Scalable Optimization of Large-scale Architecture for Reasoning
Chen Li, Yinyi Luo, Anudeep Bolimera +4
Large Language Models excel in reasoning yet often rely on Chain-of-Thought prompts, limiting performance on tasks demanding more nuanced topological structures. We present SOLAR (…