8 citations · 22 across the 13 of their papers we have counts for
15 papers
RoboTalk: Learning Multi-Robot Communication and Coordination from Multimodal Demonstrations
Dorian Benhamou Goldfajn, Mason Nakamura, Saaduddin Mahmud +3
Multi-robot collaboration could enable more efficient and scalable solutions to complex robotic tasks, but collaboration under partial observability remains challenging. Natural-la…
FrankenReport: Early Exiting in Long-Form Generation Using Expected Value of Computation
Zhengping Jiang, Gonzalo Ramos, Jina Suh +6
While deep research systems address interactive information-seeking needs impressively, their real-world deployments face latency and resource-consumption challenges. We present Fr…
RHO: Your Coding Agent is Secretly a Roboticist
Karim Elmaaroufi, Justin Svegliato, Sarunas Kalade +3
Code-as-Policies (CaP) has shown that large language models (LLMs) can write code to solve robotics tasks by composing perception, planning, and control primitives. Recent CaP syst…
Fine-Tuning LLMs with Fine-Grained Human Feedback on Text Spans
Sky CH-Wang, Justin Svegliato, Helen Appel +1
We present a method and dataset for fine-tuning language models with preference supervision using feedback-driven improvement chains. Given a model response, an annotator provides…
GRAID: Enhancing Spatial Reasoning of VLMs Through High-Fidelity Data Generation
Karim Elmaaroufi, Liheng Lai, Justin Svegliato +3
Vision Language Models (VLMs) achieve strong performance on many vision-language tasks but often struggle with spatial reasoning$\unicode{x2014}$a prerequisite for many application…
MICE for CATs: Model-Internal Confidence Estimation for Calibrating Agents with Tools
Nishant Subramani, Jason Eisner, Justin Svegliato +3
Tool-using agents that act in the world need to be both useful and safe. Well-calibrated model confidences can be used to weigh the risk versus reward of potential actions, but pri…