Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Can LLMs Reliably Self-Report Adversarial Prefills, and How?
Quang Minh Nguyen, Uzair Ahmed, Taegyoon Kim
Prior work shows that large language models (LLMs) exhibit introspective capability on benign tasks. We extend the question to safety contexts and examine how reliably a model can…
cs.CL2025
Is External Information Useful for Stance Detection with LLMs?
Quang Minh Nguyen, Taegyoon Kim
In the stance detection task, a text is classified as either favorable, opposing, or neutral towards a target. Prior work suggests that the use of external information, e.g., excer…