papers
Publications (12)
cs.LG2026
Using Reward Uncertainty to Induce Diverse Behaviour in Reinforcement Learning
Anthony GX-Chen, Ankit Anand, Gheorghe Comanici +7
cs.AI2026
An AI system to help scientists write expert-level empirical software
Eser Aygün, Anastasiya Belyaeva, Gheorghe Comanici +39
cs.LG2022
Learning how to Interact with a Complex Interface using Hierarchical Reinforcement Learning
Gheorghe Comanici, Amelia Glaese, Anita Gergely +5
cs.LG2024
Vision-Language Models as a Source of Rewards
Kate Baumli, Satinder Baveja, Feryal Behbahani +24
cs.CL2025
Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Gheorghe Comanici, Eric Bieber, Mike Schaekermann +3431
cs.LG2021
AndroidEnv: A Reinforcement Learning Platform for Android
Daniel Toyama, Philippe Hamel, Anita Gergely +6
cs.AI2024
Finding Increasingly Large Extremal Graphs with AlphaZero and Tabu Search
Abbas Mehrabian, Ankit Anand, Hyunjik Kim +16
cs.LG2021
Temporally Abstract Partial Models
Khimya Khetarpal, Zafarali Ahmed, Gheorghe Comanici +1
cs.AI2021
The Option Keyboard: Combining Skills in Reinforcement Learning
André Barreto, Diana Borsa, Shaobo Hou +8
cs.LG2026
Affordances Enable Partial World Modeling with LLMs
Khimya Khetarpal, Gheorghe Comanici, Jonathan Richens +5
cs.CL2024
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Gemini Team, Petko Georgiev, Ving Ian Lei +1132
cs.LG2020
What can I do here? A Theory of Affordances in Reinforcement Learning
Khimya Khetarpal, Zafarali Ahmed, Gheorghe Comanici +2