3 papers
cs.LG2026
Olmo Hybrid: From Theory to Practice and Back
William Merrill, Yanhong Li, Tyler Romero +19
Recent work has demonstrated the potential of non-transformer language models, especially linear recurrent neural networks (RNNs) and hybrid models that mix recurrence and attentio…
cs.LG2026
OpenJarvis: Personal AI, On Personal Devices
Jon Saad-Falcon, Avanika Narayan, Robby Manihani +10
Personal AI stacks, like OpenClaw and Hermes Agent, are becoming central to daily work, yet they route nearly every query (often over sensitive local data) to cloud-hosted frontier…
cs.LG2025
Think, Prune, Train, Improve: Scaling Reasoning without Scaling Models
Caia Costello, Simon Guo, Anna Goldie +1
Large language models (LLMs) have demonstrated strong capabilities in programming and mathematical reasoning tasks, but are constrained by limited high-quality training data. Synth…