activity
20242026
collaborators

6 papers

cs.AI2026

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight

Christopher Z. Cui, Taylor W. Killian, Prithviraj Ammanabrolu

Reasoning in Large Language Models (LLMs) poses a challenge for oversight as many misaligned behaviors do not surface until reasoning concludes. To address this, we introduce Behav…

cs.LG2026

DeltaPrompts: Escaping the Zero-Delta Trap in Multimodal Distillation

Jaehun Jung, Hyunwoo Kim, Brandon Cui +4

Distillation enables compact Vision-Language Models (VLMs) to obtain strong reasoning capabilities, yet the prompts driving this process are typically chosen via simple heuristics…

cs.LG2026

How Reasoning Evolves from Post-Training Data: An Empirical Study Using Chess

Lucas Dionisopoulos, Nicklas Majamaki, Prithviraj Ammanabrolu

We study how reasoning evolves in a language model -- from supervised fine-tuning (SFT) to reinforcement learning (RL) -- by analyzing how a set of theoretically-inspired datasets…

cs.AI2026

Beyond Needle(s) in the Embodied Haystack: Environment, Architecture, and Training Considerations for Long Context Reasoning

Bosung Kim, Prithviraj Ammanabrolu

We introduce -THOR, a new framework for long-horizon embodied tasks that advances long-context understanding in embodied AI. -THOR provides: (1) a generation framew…

cs.MA2025

Collaborating Action by Action: A Multi-agent LLM Framework for Embodied Reasoning

Isadora White, Kolby Nottingham, Ayush Maniar +5

Collaboration is ubiquitous and essential in day-to-day life -- from exchanging ideas, to delegating tasks, to generating plans together. This work studies how LLMs can adaptively…

cs.HC2024

CPS-TaskForge: Generating Collaborative Problem Solving Environments for Diverse Communication Tasks

Nikita Haduong, Irene Wang, Bo-Ru Lu +2

Teams can outperform individuals; could adding AI teammates further bolster performance of teams solving problems collaboratively? Collaborative problem solving (CPS) research comm…