3 citations · 3 across the 4 of their papers we have counts for
Showing 2026Show all
2 papers · 1 filter
cs.AI2026
AI Research Preference Models
Thomas Simon Foster, Bassel Al Omari, Tingchen Fu +30
AI research agents (AIRA) can now carry machine learning experiments from proposal through implementation and evaluation. Yet progress on frontier tasks is throttled by the cost of…
cs.LG2026
Finding the Time to Think: Learning Planning Budgets in Real-Time RL
Aneesh Muppidi, Firas Darwish, Dylan Cope +2
Deliberating takes time. In real-time settings, that time is not free. Standard reinforcement learning (RL) sidesteps this as the environment waits indefinitely for the agent's dec…