Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
Uncovering the Computational Ingredients of Human-Like Representations in LLMs
Zach Studdiford, Timothy T. Rogers, Kushin Mukherjee +1
The human ability to translate diverse perceptual and linguistic inputs into structured behavior has been thought to rest on learning robust representations of concepts. The rapid…
cs.AI2025
Evaluating Steering Techniques using Human Similarity Judgments
Zach Studdiford, Timothy T. Rogers, Siddharth Suresh +1
Current evaluations of Large Language Model (LLM) steering techniques focus on task-specific performance, overlooking how well steered representations align with human cognition. U…