5 papers · 1 filter
Agentic Systems as Boosting Weak Reasoning Models
Varun Sunkaraneni, Pierfrancesco Beneventano, Riccardo Neumarker +2
Can a committee of weak reasoning-model calls reach the performance of much stronger models? We study verifier-backed committee search as inference-time boosting for reasoning lang…
The Generalized Turing Test: A Foundation for Comparing Intelligence
Daniel Mitropolsky, Susan S. Hong, Riccardo Neumarker +2
We introduce the Generalized Turing Test (GTT), a formal framework for comparing the capabilities of arbitrary agents via indistinguishability. For agents A and B, we define the Tu…
pAI/MSc: ML Theory Research with Humans on the Loop
Mahmoud Abdelmoneum, Pierfrancesco Beneventano, Tomaso Poggio
We present pAI/MSc, an open-source, customizable, modular multi-agent system for academic research workflows. Our goal is not autonomous scientific ideation, nor fully automated re…
Tool Building as a Path to "Superintelligence"
David Koplow, Tomer Galanti, Tomaso Poggio
The Diligent Learner framework suggests LLMs can achieve superintelligence via test-time search, provided a sufficient step-success probability . In this work, we design a benc…
What if Eye...? Computationally Recreating Vision Evolution
Kushagra Tiwary, Aaron Young, Zaid Tasneem +6
Vision systems in nature show remarkable diversity, from simple light-sensitive patches to complex camera eyes with lenses. While natural selection has produced these eyes through…