collaborators

6 papers

cs.LG2026

Learning Multi-Agent Coordination via Sheaf-ADMM

Jeffrey Seely, Bartłomiej Cupiał, Llion Jones

We present a differentiable optimization framework for multi-agent coordination. An input is decomposed into overlapping local views, each processed by an agent that solves a conve…

cs.CL2026

Fast-weight Product Key Memory

Tianyu Zhao, Llion Jones

Sequence modeling layers in modern language models typically face a trade-off between storage capacity and computational efficiency. While softmax attention offers unbounded storag…

cs.LG2025

Continuous Thought Machines

Luke Darlow, Ciaran Regan, Sebastian Risi +2

Biological brains demonstrate complex neural activity, where neural dynamics are critical to how brains process information. Most artificial neural networks ignore the complexity o…

cs.CL2025

Building Tailored Speech Recognizers for Japanese Speaking Assessment

Yotaro Kubo, Richard Sproat, Chihiro Taguchi +1

This paper presents methods for building speech recognizers tailored for Japanese speaking assessment tasks. Specifically, we build a speech recognizer that outputs phonemic labels…

cs.CL2025

TransEvalnia: Reasoning-based Evaluation and Ranking of Translations

Richard Sproat, Tianyu Zhao, Llion Jones

We present TransEvalnia, a prompting-based translation evaluation and ranking system that uses reasoning in performing its evaluations and ranking. This system presents fine-graine…

cs.AI2025

Sudoku-Bench: Evaluating creative reasoning with Sudoku variants

Jeffrey Seely, Yuki Imajuku, Tianyu Zhao +2

Existing reasoning benchmarks for large language models (LLMs) frequently fail to capture authentic creativity, often rewarding memorization of previously observed patterns. We add…