Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight
Christopher Z. Cui, Taylor W. Killian, Prithviraj Ammanabrolu
Reasoning in Large Language Models (LLMs) poses a challenge for oversight as many misaligned behaviors do not surface until reasoning concludes. To address this, we introduce Behav…
cs.AI2026
Beyond Needle(s) in the Embodied Haystack: Environment, Architecture, and Training Considerations for Long Context Reasoning
Bosung Kim, Prithviraj Ammanabrolu
We introduce -THOR, a new framework for long-horizon embodied tasks that advances long-context understanding in embodied AI. -THOR provides: (1) a generation framew…