collaborators

7 papers

cs.DB2026

Evergreen: Efficient Claim Verification for Semantic Aggregates

Alexander W. Lee, Benjamin Han, Shayak Sen +3

With recent semantic query processing engines, semantic aggregation has become a primitive operator, enabling the reduction of a relation into a natural language aggregate using an…

cs.AI2026

Plans Don't Persist: Why Context Management Is Load Bearing for LLM Agents

Aman Mehta, Anupam Datta

Long-horizon agents depend on context management: systems compress, summarize, and evict old tokens so tasks can continue beyond finite windows. That is safe only when dropped info…

cs.DB2026

Larch: Learned Query Optimization for Semantic Predicates

Fuheng Zhao, Pawel Liskowski, Zihan Li +5

With the advent of Large Language Models (LLMs), many database systems introduced semantic operators that enabled analytical queries over unstructured data (e.g. text, images, vide…

cs.DB2026

AvalancheBench: Evaluating Enterprise Data Agents Through Latent World Recovery

Darek Kleczek, Fuheng Zhao, Alexander W. Lee +4

We introduce AvalancheBench, a benchmark for evaluating enterprise data agents through \emph{latent world recovery}. AvalancheBench improves on existing benchmarks in three ways. F…

cs.DB2026

Cortex AISQL: A Production SQL Engine for Unstructured Data

Paweł Liskowski, Benjamin Han, Paritosh Aggarwal +11

Snowflake's Cortex AISQL is a production SQL engine that integrates native semantic operations directly into SQL. This integration allows users to write declarative queries that co…

cs.CL2026

Strategic Navigation or Stochastic Search? How Agents and Humans Reason Over Document Collections

Łukasz Borchmann, Jordy Van Landeghem, Michał Turski +12

Multimodal agents offer a promising path to automating complex document-intensive workflows. Yet, a critical question remains: do these agents demonstrate genuine strategic reasoni…