◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Ming Kong

11 papers hereh-index 9303 citations35 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author10

Across the 10 of 11 papers where every author was matched, so the position is known.

fields
  • cs.CV6
  • cs.AI3
  • cs.GR1
  • cs.MA1
ORCID 0000-0002-6177-3707

identity via Semantic Scholar / OpenAlex

works on
feedback-driven adaptation 1interactive agents 1large language models 1procedural memory 1task automation 1

From the 1 of 11 linked papers with an AI index.

collaborators
Showing cs.AIShow all

4 papers · 1 filter

cs.AI2026

MetaRAG: Belief-Action Aligned Policy Optimization for Agentic RAG

Qiuyi Qi, Tian Liang, Jiamu Wang +7

Agentic retrieval-augmented generation (RAG) requires language models to decide when to continue searching and when to answer. Existing RL-based methods rely on external supervisio…

cs.AI2026

D2R2: Discrete Diffusion with Regulation Reinforcement for Single-Cell Perturbation Prediction

Ninghan Fan, Qi Liu, Xunuo Zhu +9

Predicting single-cell transcriptomic responses to genetic perturbations is central to functional genomics and virtual-cell modeling. Existing approaches, however, typically predic…

cs.AI2026

STAPO: Selective Trajectory-Aware Policy Optimization for LLM Agent Training

Qiuyi Qi, Tian Liang, Mutian Bao +8

Reinforcement Learning (RL) is the dominant paradigm for training Large Language Model (LLM) agents on long-horizon tasks. However, sparse and delayed rewards often lead to traject…

cs.AI2026

CARL: Constraint-Aware Reinforcement Learning for Planning with LLMs

Qiuyi Qi, Jinjian Zhang, Mutian Bao +9

Despite their strong reasoning capabilities and extensive world knowledge, Large Language Models (LLMs) frequently generate plans that violate task constraints, undermining their r…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.