Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
LaneRoPE: Positional Encoding for Collaborative Parallel Reasoning and Generation
Gabriele Cesa, Thomas Hehn, Aleix Torres-Camps +4
Parallel LLM test-time scaling techniques (e.g., best-of-) require drawing sequences conditioned on the same input prompt. These methods boost accuracy while exploiting th…
cs.AI2025
Local Look-Ahead Guidance via Verifier-in-the-Loop for Automated Theorem Proving
Sara Rajaee, Kumar Pratik, Gabriele Cesa +1
The most promising recent methods for AI reasoning require applying variants of reinforcement learning (RL) either on rolled out trajectories from the LLMs, even for the step-wise…