collaborators

6 papers

cs.AI2026

FlavourBench: Executable Culinary Reward Maps for Language Model Evaluation and Post-Training

Josef Chen, Erim Hayretci

We introduce FlavorBench: a benchmark for Compiling Dense Deterministic Answer Maps from a Versioned Culinary Embeddings Model. We test 27 frontier large language model endpoints o…

cs.DC2026

Ready Cohorts: Bounding GPU Opportunity and Avoiding Host Round Trips in LLM-Agent Control

Josef Liyanjun Chen

LLM-agent services repeatedly execute small deterministic transitions between model and tool calls: route an outcome, update state, and emit the next effect. We ask when this contr…

cs.AI2026

When Does Combining Language Models Help? A Co-Failure Ceiling on Routing, Voting, and Mixture-of-Agents Across 67 Frontier Models

Josef Chen

Multi-model LLM systems such as routing, voting, cascades, fusion, and mixture-of-agents are used to beat single-model accuracy. We show that their gain is capped by a quantity the…

cs.AI2026

Memory as a Wasting Asset: Pricing Flash Endurance for Embodied Agents, and the Limits of Doing So

Josef Liyanjun Chen

A robot's flash endurance is a non-renewable stock: every persisted write spends one of a few thousand program/erase cycles and never refills, yet no fielded robot memory system pr…

cs.AI2026

Epicure: Navigating the Emergent Geometry of Food Ingredient Embeddings

Jakub Radzikowski, Josef Chen

We present Epicure, a family of three sibling skip-gram ingredient embeddings retrained from scratch on a multilingual recipe corpus. We aggregate 4.14M recipes from 11 sources spa…

cs.CY2026

Epicure: Multidimensional Flavor Structure in Food Ingredient Embeddings

Jakub Radzikowski, Josef Chen

A chef's intuition about flavor, texture, and cultural identity represents tacit knowledge that is difficult to articulate yet central to culinary practice. We show that this knowl…