activity
20202026
most citedTensor Trust: Interpretable Prompt Injection Attacks from an Online Game

8 citations · 22 across the 13 of their papers we have counts for

collaborators

15 papers

cs.RO2026

RoboTalk: Learning Multi-Robot Communication and Coordination from Multimodal Demonstrations

Dorian Benhamou Goldfajn, Mason Nakamura, Saaduddin Mahmud +3

Multi-robot collaboration could enable more efficient and scalable solutions to complex robotic tasks, but collaboration under partial observability remains challenging. Natural-la…

cs.HC2026

FrankenReport: Early Exiting in Long-Form Generation Using Expected Value of Computation

Zhengping Jiang, Gonzalo Ramos, Jina Suh +6

While deep research systems address interactive information-seeking needs impressively, their real-world deployments face latency and resource-consumption challenges. We present Fr…

cs.RO2026

RHO: Your Coding Agent is Secretly a Roboticist

Karim Elmaaroufi, Justin Svegliato, Sarunas Kalade +3

Code-as-Policies (CaP) has shown that large language models (LLMs) can write code to solve robotics tasks by composing perception, planning, and control primitives. Recent CaP syst…

cs.CL2025

Fine-Tuning LLMs with Fine-Grained Human Feedback on Text Spans

Sky CH-Wang, Justin Svegliato, Helen Appel +1

We present a method and dataset for fine-tuning language models with preference supervision using feedback-driven improvement chains. Given a model response, an annotator provides…

cs.CV2025

GRAID: Enhancing Spatial Reasoning of VLMs Through High-Fidelity Data Generation

Karim Elmaaroufi, Liheng Lai, Justin Svegliato +3

Vision Language Models (VLMs) achieve strong performance on many vision-language tasks but often struggle with spatial reasoning$\unicode{x2014}$a prerequisite for many application…

cs.CL2025

MICE for CATs: Model-Internal Confidence Estimation for Calibrating Agents with Tools

Nishant Subramani, Jason Eisner, Justin Svegliato +3

Tool-using agents that act in the world need to be both useful and safe. Well-calibrated model confidences can be used to weigh the risk versus reward of potential actions, but pri…