2 papers
cs.AI2026
Detecting the Disturbance: A Nuanced View of Introspective Abilities in LLMs
Ely Hahami, Ishaan Sinha, Lavik Jain +2
Can large language models introspect, that is, accurately detect perturbations to their own internal states? We systematically investigate this question using activation steering i…
cs.CL2024
GPT-4o System Card
OpenAI, :, Aaron Hurst +416
GPT-4o is an autoregressive omni model that accepts as input any combination of text, audio, image, and video, and generates any combination of text, audio, and image outputs. It's…