1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.AI2025★ 1 cited
Toward a Theory of Generalizability in LLM Mechanistic Interpretability Research
Sean Trott
Research on Large Language Models (LLMs) increasingly focuses on identifying mechanistic explanations for their behaviors, yet the field lacks clear principles for determining when…
cs.CV2025
Seeing Through Words, Speaking Through Pixels: Deep Representational Alignment Between Vision and Language Models
Zoe Wanying He, Sean Trott, Meenakshi Khosla
Recent studies show that deep vision-only and language-only models--trained on disjoint modalities--nonetheless project their inputs into a partially aligned representational space…