◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Mehdi Mounsif

2 papers hereh-index 216 citations4 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • last author2

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.LG2

identity via Semantic Scholar / OpenAlex

collaborators

2 papers

cs.LG2025

Large Reasoning Models are not thinking straight: on the unreliability of thinking trajectories

Jhouben Cuesta-Ramirez, Samuel Beaussant, Mehdi Mounsif

Large Language Models (LLMs) trained via Reinforcement Learning (RL) have recently achieved impressive results on reasoning benchmarks. Yet, growing evidence shows that these model…

cs.LG2025

Scaling Algorithm Distillation for Continuous Control with Mamba

Samuel Beaussant, Mehdi Mounsif

Algorithm Distillation (AD) was recently proposed as a new approach to perform In-Context Reinforcement Learning (ICRL) by modeling across-episodic training histories autoregressiv…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.