6 papers
What Drives Interactive Improvement from Feedback?
Bartłomiej Cupiał, Jan Łojek, Mikołaj Garstecki +3
We study when natural-language feedback produces improvement beyond the gains obtainable from repeated attempts alone. In multi-turn language agent setting, higher final accuracy c…
When Does Non-Uniform Replay Matter in Reinforcement Learning?
Michal Korniak, Mikołaj Czarnecki, Yarden As +3
Modern off-policy reinforcement learning algorithms often rely on simple uniform replay sampling and it remains unclear when and why non-uniform replay improves over this strong ba…
RapidDock: Unlocking Proteome-scale Molecular Docking
Rafał Powalski, Bazyli Klockiewicz, Maciej Jaśkowski +6
Accelerating molecular docking -- the process of predicting how molecules bind to protein targets -- could boost small-molecule drug discovery and revolutionize medicine. Unfortuna…
Repurposing Language Models into Embedding Models: Finding the Compute-Optimal Recipe
Alicja Ziarko, Albert Q. Jiang, Bartosz Piotrowski +3
Text embeddings are essential for many tasks, such as document retrieval, clustering, and semantic similarity assessment. In this paper, we study how to contrastively train text em…
What Matters in Hierarchical Search for Combinatorial Reasoning Problems?
Michał Zawalski, Gracjan Góral, Michał Tyrolski +5
Efficiently tackling combinatorial reasoning problems, particularly the notorious NP-hard tasks, remains a significant challenge for AI research. Recent efforts have sought to enha…
Bigger, Regularized, Optimistic: scaling for compute and sample-efficient continuous control
Michal Nauman, Mateusz Ostaszewski, Krzysztof Jankowski +2
Sample efficiency in Reinforcement Learning (RL) has traditionally been driven by algorithmic enhancements. In this work, we demonstrate that scaling can also lead to substantial i…