◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Tong Che

5 papers hereh-index 3131 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author2
  • last author1

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.AI3
  • cs.LG2
same name
  • Tong Che — 6 papers, h 3
  • Tong Che — 2 papers, h 3
  • Tong Che — 2 papers, h 6
  • Tong Che — 1 paper, h 0

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

5 papers

cs.LG2026

Learning with a Single Rollout via Monte Carlo Pass@k Critic

Fengdi Che, Yang Liu, Lei Yu +4

Estimating token-level advantages in reinforcement learning (RL) for language models remains challenging because scaling up episodic experience collection is expensive. The difficu…

cs.AI2026

Greed Is Learned: Visible Incentives as Reward-Hacking Triggers

Tong Che, Rui Wu

Deployed agents increasingly act with their reward proxy in view, such as a balance, score, or KPI dashboard. We show that reinforcement learning can make a policy \emph{addicted}…

cs.LG2026

Constitutional Value Potentials: reading and steering internal priority margins in language models

Tong Che, Rui Wu

A constitution tells a language model what to value, but little tells us whether it does. Adherence is judged from outputs, and output evidence is most fragile on value conflicts,…

cs.AI2026

Reference Feature Atlases for Mechanistic Auditing of Language Models

Rui Wu, Tong Che

Auditing a new language model usually means relearning and reinterpreting its internal features from scratch. We propose a reference feature atlas: a sparse feature library trained…

cs.AI2024

LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Di Zhang, Jianbo Wu, Jingdi Lei +9

This paper presents an advanced mathematical problem-solving framework, LLaMA-Berry, for enhancing the mathematical reasoning ability of Large Language Models (LLMs). The framework…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.