◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Min Cai

7 papers hereh-index 5123 citations15 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author6

Across the 7 of 7 papers where every author was matched, so the position is known.

fields
  • cs.CL4
  • cs.LG2
  • cs.AI1
same name
  • Min Cai — 2 papers, h 6
  • Min Cai — 2 papers
  • Min Cai — 1 paper
  • Min Cai — 1 paper
  • Min Cai — 1 paper
  • Min Cai — 1 paper, h 1

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedTDRM: Smooth Reward Models with Temporal Difference for LLM RL and Inference

1 citations · 1 across the 4 of their papers we have counts for

collaborators
Showing cs.CLShow all

4 papers · 1 filter

cs.CL2026

Advancing General-Purpose Reasoning Models with Modular Gradient Surgery

Min Cai, Yu Liang, Longzheng Wang +6

Reinforcement learning (RL) has played a central role in recent advances in large reasoning models (LRMs), yielding strong gains in verifiable and open-ended reasoning. However, tr…

cs.CL2026

Talk Less, Verify More: Improving LLM Assistants with Semantic Checks and Execution Feedback

Yan Sun, Ming Cai, Stanley Kok

As large language model (LLM) assistants become increasingly integrated into enterprise workflows, their ability to generate accurate, semantically aligned, and executable outputs…

cs.CL2025

How Post-Training Reshapes LLMs: A Mechanistic View on Knowledge, Truthfulness, Refusal, and Confidence

Hongzhe Du, Weikai Li, Min Cai +5

Post-training is essential for the success of large language models (LLMs), transforming pre-trained base models into more useful and aligned post-trained models. While plenty of w…

cs.CL2025

DataSciBench: An LLM Agent Benchmark for Data Science

Dan Zhang, Sining Zhoubian, Min Cai +7

This paper presents DataSciBench, a comprehensive benchmark for evaluating Large Language Model (LLM) capabilities in data science. Recent related benchmarks have primarily focused…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.