2 citations · 8 across the 15 of their papers we have counts for
17 papers
COvolve: Adversarial Co-Evolution of Large-Language-Model-Generated Policies and Environments via Two-Player Zero-Sum Game
Alkis Sygkounas, Rishi Hazra, Andreas Persson +2
A central challenge in building continually improving agents is that training environments are typically static or manually constructed. This restricts continual learning and gener…
APC-RL: Exceeding Data-Driven Behavior Priors with Adaptive Policy Composition
Finn Rietz, Pedro Zuidberg dos Martires, Johannes Andreas Stork
Incorporating demonstration data into reinforcement learning (RL) can greatly accelerate learning, but existing approaches often assume demonstrations are optimal and fully aligned…
LexiCon: a Benchmark for Planning under Temporal Constraints in Natural Language
Periklis Mantenoglou, Rishi Hazra, Pedro Zuidberg Dos Martires +1
Owing to their reasoning capabilities, large language models (LLMs) have been evaluated on planning tasks described in natural language. However, LLMs have largely been tested on p…
A Quantum Information Theoretic Approach to Tractable Probabilistic Models
Pedro Zuidberg Dos Martires
By recursively nesting sums and products, probabilistic circuits have emerged in recent years as an attractive class of generative models as they enjoy, for instance, polytime marg…
Have Large Language Models Learned to Reason? A Characterization via 3-SAT Phase Transition
Rishi Hazra, Gabriele Venturato, Pedro Zuidberg Dos Martires +1
Large Language Models (LLMs) have been touted as AI models possessing advanced reasoning abilities. In theory, autoregressive LLMs with Chain-of-Thought (CoT) can perform more seri…
Automated Reasoning in Systems Biology: a Necessity for Precision Medicine
Pedro Zuidberg Dos Martires, Vincent Derkinderen, Luc De Raedt +1
Recent developments in AI have reinvigorated pursuits to advance the (life) sciences using AI techniques, thereby creating a renewed opportunity to bridge different fields and find…