7 citations · 8 across the 4 of their papers we have counts for
4 papers
Code as Reward: Empowering Reinforcement Learning with VLMs
David Venuto, Sami Nur Islam, Martin Klissarov +3
Pre-trained Vision-Language Models (VLMs) are able to understand visual concepts, describe and decompose complex tasks into sub-tasks, and provide feedback on task completion. In t…
Transfer Learning for the Prediction of Entity Modifiers in Clinical Text: Application to Opioid Use Disorder Case Detection
Abdullateef I. Almudaifer, Whitney Covington, JaMor Hairston +9
Background: The semantics of entities extracted from a clinical text can be dramatically altered by modifiers, including entity negation, uncertainty, conditionality, severity, and…
Accelerating exploration and representation learning with offline pre-training
Bogdan Mazoure, Jake Bruce, Doina Precup +2
Sequential decision-making agents struggle with long horizon tasks, since solving them requires multi-step reasoning. Most reinforcement learning (RL) algorithms address this chall…
Proving Theorems using Incremental Learning and Hindsight Experience Replay
Eser Aygün, Laurent Orseau, Ankit Anand +5
Traditional automated theorem provers for first-order logic depend on speed-optimized search and many handcrafted heuristics that are designed to work best over a wide range of dom…