1 citations · 1 across the 17 of their papers we have counts for
43 papers
GlossoGen: Emergent Language in Complex Multi-Agent LLM Interactions
Elias Stengel-Eskin, Newton Sander, Carlos Bonetti +4
The growing rate at which LLM agents interact with one another raises key questions about language evolution in multi-LLM-agent settings, with implications for safety and monitorab…
Who is the Agent to Blame? Localizing Faithfulness and Citation Mistakes in Agentic Deep Research
Eran Hirsch, David Wan, Han Wang +3
Deep research (DR) systems produce long-form cited reports by orchestrating multiple agents that search and synthesize information from the web. Citations are the primary mechanism…
CreativeInstruct: Scalably Teaching LLMs to Balance Quality, Creativity, and Diversity
Ananya Sahu, Mohit Bansal, Elias Stengel-Eskin
While post-training improves the capabilities of large language models (LLMs), it generally lowers their output diversity and creativity, negatively impacting tasks that explicitly…
Physics Question Scene Graph: Fine-grained Evaluation of Physical Plausibility in Text-to-Video Generation
Atin Pothiraj, Jaemin Cho, Yue Zhang +2
Video generation models are increasingly capable of producing realistic videos, but they still struggle to generate videos that follow basic physical laws. Compounding this is a la…
CalVerT: Augmenting Agents with Calibrated Verifier Telemetry Improves Action and Learning in Knowledge-Intensive Tasks
Ashwin Vinod, Ying Ding, Elias Stengel-Eskin
LLM agents in knowledge intensive question answering take retrieval and reasoning actions with incomplete knowledge about whether their current answer is uncertain, unsupported, or…
PragReST: Self-Reinforcing Counterfactual Reasoning for Pragmatic Language Understanding
Jihyung Park, Minchao Huang, Leqi Liu +1
Natural language understanding often depends on meanings that are implied rather than explicitly stated, requiring pragmatic reasoning. Despite strong performance on math and logic…