works on

From the 1 of 29 linked papers with an AI index.

collaborators

29 papers

cs.LG2026

DHRCL:Training Code LLMs with Dense Hierarchical Rewards and Curriculum Learning

Shuhang Wang, Ziming Li, Hui Cheng

The paper introduces DHRCL, a reinforcement‑learning framework for code‑focused large language models that uses a hierarchy of dense rewards (syntax, execution, unit‑test pass, and…

cs.CL2026

Kwai Summary Attention Technical Report

Chenglong Chu, Guorui Zhou, Guowang Zhang +35

Long-context ability, has become one of the most important iteration direction of next-generation Large Language Models, particularly in semantic understanding/reasoning, code agen…

cs.CV2026

UniPR-3D: Towards Universal Visual Place Recognition with Visual Geometry Grounded Transformer

Tianchen Deng, Xun Chen, Ziming Li +5

Visual Place Recognition (VPR) has been traditionally formulated as a single-image retrieval task. Using multiple views offers clear advantages, yet this setting remains relatively…

cs.LG2026

SupraBench: A Benchmark for Supramolecular Chemistry

Tianyi Ma, Yijun Ma, Zehong Wang +6

Supramolecular chemistry, which includes the study of non-covalent host-guest assemblies, has advanced various applications. However, designing host-guest systems remains time-cons…

cs.AI2026

MDForge: Agentic Molecular Dynamics Pipeline Design under Sparse Simulator Feedback

Zehong Wang, Yijun Ma, Connor R. Schmidt +7

Molecular dynamics (MD) is the canonical in-silico method for atomistic molecular science, simulating molecular behavior from first-principle physics. Designing an MD pipeline for…

cs.LG2026

ProPlay: Procedural World Models for Self-Evolving LLM Agents

Yijun Ma, Zehong Wang, Yiyang Li +5

Self-evolving agents are expected to improve through interaction without external supervision, but this remains difficult in partially observable environments where agents must exp…