works on

From the 1 of 10 linked papers with an AI index.

collaborators

10 papers

cs.AI2026

Tycho: Active Abstraction with Programmatic World Models for ARC-AGI-3

Jens Lehmann, Andrei Aioanei, Sahar Vahdati

The paper presents Tycho, a coding-agent that actively builds and uses programmatic world models to infer game rules and improve action efficiency in the ARC-AGI-3 benchmark, achie…

cs.AI2026

Graphlets as Building Blocks for Structural Vocabulary in Knowledge Graph Foundation Models

Kossi Amouzouvi, Robert Wardenga, Jens Lehmann +1

Foundation models excel at language, where sentences become tokens, and vision, where images become pixels, because both reduce to discrete symbols on a shared, fixed grid. Knowled…

cs.IR2026

Prompt Compression in the Wild: Measuring Latency, Rate Adherence, and Quality for Faster LLM Inference

Cornelius Kummer, Lena Jurkschat, Michael Färber +1

With the wide adoption of language models for IR -- and specifically RAG systems -- the latency of the underlying LLM becomes a crucial bottleneck, since the long contexts of retri…

cs.AI2026

The ARC of Progress towards AGI: A Living Survey of Abstraction and Reasoning

Sahar Vahdati, Andrei Aioanei, Haridhra Suresh +1

The Abstraction and Reasoning Corpus (ARC-AGI) has become a key benchmark for fluid intelligence in AI. This survey presents the first cross-generation analysis of 82 approaches ac…

cs.CL2026

ARC-TGI: Human-Validated Task Generators with Reasoning Chain Templates for ARC-AGI

Jens Lehmann, Syeda Khushbakht, Nikoo Salehfard +4

The Abstraction and Reasoning Corpus (ARC-AGI) probes few-shot abstraction and rule induction on small visual grids, but progress is difficult to measure on static collections of h…

cs.LG2026

LLM Reasoning with Process Rewards for Outcome-Guided Steps

Mohammad Rezaei, Jens Lehmann, Sahar Vahdati

Mathematical reasoning in large language models has improved substantially with reinforcement learning using verifiable rewards, where final answers can be checked automatically an…