◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Ousmane Amadou Dia

3 papers hereh-index 5890 citations20 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author2
  • middle author1

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • stat.ML2
  • cs.CL1

identity via Semantic Scholar / OpenAlex

collaborators

3 papers

stat.ML2026

Adaptive Nucleus Truncation for Long-Form Reasoning

Ousmane Amadou Dia

Sampling plays an important role in long-form language-model reasoning. Over thousands of decoding steps, small changes in the candidate token set can compound into different reaso…

stat.ML2026

Variational Proximal Policy Optimization

Ousmane Amadou Dia

Reinforcement Learning from Human Feedback via Proximal Policy Optimization often suffers from policy mode collapse, brittle exploration loops, and distribution drift. This paper i…

cs.CL2026

LH-Deception: Simulating and Understanding LLM Deceptive Behaviors in Long-Horizon Interactions

Yang Xu, Xuanming Zhang, Samuel Yeh +4

Deception is a pervasive feature of human communication and an emerging concern in large language models (LLMs). While recent studies document instances of LLM deception, most eval…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.