◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Andrea Zanette

5 papers hereh-index 5513 citations6 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • last author5

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.LG4
  • cs.CL1
same name
  • Andrea Zanette — 4 papers, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

5 papers

cs.CL2026

Message Passing Enables Efficient Reasoning

Xuecheng Liu, Daman Arora, Gokul Swamy +1

While inference-time scaling has improved the reasoning abilities of large language models (LLMs), the need to generate long chains-of-thought (CoTs) is a computational bottleneck.…

cs.LG2026

Expanding the Capabilities of Reinforcement Learning via Text Feedback

Yuda Song, Lili Chen, Fahim Tajwar +5

The success of RL for LLM post-training stems from an unreasonably uninformative source: a single bit of information per rollout as binary reward or preference label. At the other…

cs.LG2026

Maximum Likelihood Reinforcement Learning

Fahim Tajwar, Guanning Zeng, Yueer Zhou +7

Reinforcement learning (RL) is the method of choice for training models in setups where the objective function can only be evaluated by sampling from the model. Our key observation…

cs.LG2025

Training Language Models to Reason Efficiently

Daman Arora, Andrea Zanette

Scaling model size and training data has led to great advances in the performance of Large Language Models (LLMs). However, the diminishing returns of this approach necessitate alt…

cs.LG2025

Can Large Reasoning Models Self-Train?

Sheikh Shafayat, Fahim Tajwar, Ruslan Salakhutdinov +2

Recent successes of reinforcement learning (RL) in training large reasoning models motivate the question of whether self-training - the process where a model learns from its own ju…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.