5 papers
GeoRA: Geometry-Aware Low-Rank Adaptation for RLVR
Jiaying Zhang, Lei Shi, Jiguo Li +4
Reinforcement Learning with Verifiable Rewards (RLVR) is a key paradigm for improving large-scale reasoning models. Unlike supervised fine-tuning (SFT), RLVR exhibits distinct opti…
Acceleration via Perturbations on Low-resolution Ordinary Differential Equations
Xudong Li, Lei Shi, Mingqi Song
Recently, the high-resolution ordinary differential equation (ODE) framework, which retains higher-order terms, has been proposed to analyze gradient-based optimization algorithms.…
Dynamic Water-Wave Tweezers
Jun Wang, Shanhe Pang, Zhiyuan Che +8
Following a recent demonstration of stable trapping of floating particles by stationary (monochromatic) structured water waves [Nature 638, 394 (2025)], we report dynamic water-wav…
Leveraging semantic similarity for experimentation with AI-generated treatments
Lei Shi, David Arbour, Raghavendra Addanki +2
Large Language Models (LLMs) enable a new form of digital experimentation where treatments combine human and model-generated content in increasingly sophisticated ways. The main me…
Revisit First-order Methods for Geodesically Convex Optimization
Yunlu Shu, Jiaxin Jiang, Lei Shi +1
In a seminal work of Zhang and Sra, gradient descent methods for geodesically convex optimization were comprehensively studied. In particular, Zhang and Sra derived a comparison in…