activity
20242026
most citedCLaMR: Contextualized Late-Interaction for Multimodal Content Retrieval

1 citations · 1 across the 17 of their papers we have counts for

collaborators

43 papers

cs.CL2026

GlossoGen: Emergent Language in Complex Multi-Agent LLM Interactions

Elias Stengel-Eskin, Newton Sander, Carlos Bonetti +4

The growing rate at which LLM agents interact with one another raises key questions about language evolution in multi-LLM-agent settings, with implications for safety and monitorab…

cs.CL2026

Who is the Agent to Blame? Localizing Faithfulness and Citation Mistakes in Agentic Deep Research

Eran Hirsch, David Wan, Han Wang +3

Deep research (DR) systems produce long-form cited reports by orchestrating multiple agents that search and synthesize information from the web. Citations are the primary mechanism…

cs.CL2026

CreativeInstruct: Scalably Teaching LLMs to Balance Quality, Creativity, and Diversity

Ananya Sahu, Mohit Bansal, Elias Stengel-Eskin

While post-training improves the capabilities of large language models (LLMs), it generally lowers their output diversity and creativity, negatively impacting tasks that explicitly…

cs.CV2026

Physics Question Scene Graph: Fine-grained Evaluation of Physical Plausibility in Text-to-Video Generation

Atin Pothiraj, Jaemin Cho, Yue Zhang +2

Video generation models are increasingly capable of producing realistic videos, but they still struggle to generate videos that follow basic physical laws. Compounding this is a la…

cs.CL2026

CalVerT: Augmenting Agents with Calibrated Verifier Telemetry Improves Action and Learning in Knowledge-Intensive Tasks

Ashwin Vinod, Ying Ding, Elias Stengel-Eskin

LLM agents in knowledge intensive question answering take retrieval and reasoning actions with incomplete knowledge about whether their current answer is uncertain, unsupported, or…

cs.CL2026

PragReST: Self-Reinforcing Counterfactual Reasoning for Pragmatic Language Understanding

Jihyung Park, Minchao Huang, Leqi Liu +1

Natural language understanding often depends on meanings that are implied rather than explicitly stated, requiring pragmatic reasoning. Despite strong performance on math and logic…