3 papers
cs.LG2026
KV-Fold: One-Step KV-Cache Recurrence for Long-Context Inference
Alireza Nadali, Patrick Cooper, Ashutosh Trivedi +1
We introduce KV-Fold, a simple, training-free long-context inference protocol that treats the key-value (KV) cache as the accumulator in a left fold over sequence chunks. At each s…
cs.AI2026
Average Reward Reinforcement Learning for Omega-Regular and Mean-Payoff Objectives
Milad Kazemi, Mateo Perez, Fabio Somenzi +3
Recent advances in reinforcement learning (RL) have renewed interest in reward design for shaping agent behavior, but manually crafting reward functions is tedious and error-prone.…
cs.FL2024
LLMs as Probabilistic Minimally Adequate Teachers for DFA Learning
Lekai Chen, Ashutosh Trivedi, Alvaro Velasquez
The emergence of intelligence in large language models (LLMs) has inspired investigations into their integration into automata learning. This paper introduces the probabilistic Min…