3 papers
cs.AI2026
LoCA: Forward-Only LLM Tuning after One-Shot Calibration with Local Credit Assignment
Linhan Xia, Rui Liu, Zhaofeng Zhang +3
Parameter-efficient post-training reduces the number of trainable parameters, but still requires repeated end-to-end backpropagation through the frozen backbone. Every adaptation s…
cs.AI2026
Evolving in the Agent Jungle via History-Informed Opponent Awareness
Zhaofeng Zhang, Linhan Xia, Rui Liu +3
Learning to adapt strategies through interaction is a key step toward more general and autonomous LLM agents. Existing approaches typically achieve behavioral adaptation by revisin…
cs.LG2025
Adaptive Dueling Double Deep Q-networks in Uniswap V3 Replication and Extension with Mamba
Zhaofeng Zhang
The report goes through the main steps of replicating and improving the article "Adaptive Liquidity Provision in Uniswap V3 with Deep Reinforcement Learning." The replication part…