collaborators

5 papers

cs.AI2026

LoCA: Forward-Only LLM Tuning after One-Shot Calibration with Local Credit Assignment

Linhan Xia, Rui Liu, Zhaofeng Zhang +3

Parameter-efficient post-training reduces the number of trainable parameters, but still requires repeated end-to-end backpropagation through the frozen backbone. Every adaptation s…

cs.AI2026

Evolving in the Agent Jungle via History-Informed Opponent Awareness

Zhaofeng Zhang, Linhan Xia, Rui Liu +3

Learning to adapt strategies through interaction is a key step toward more general and autonomous LLM agents. Existing approaches typically achieve behavioral adaptation by revisin…

cs.LG2025

Adaptive Dueling Double Deep Q-networks in Uniswap V3 Replication and Extension with Mamba

Zhaofeng Zhang

The report goes through the main steps of replicating and improving the article "Adaptive Liquidity Provision in Uniswap V3 with Deep Reinforcement Learning." The replication part…

q-fin.MF2025

Quantformer: from attention to profit with a quantitative transformer trading strategy

Zhaofeng Zhang, Banghao Chen, Shengxin Zhu +1

In traditional quantitative trading practice, navigating the complicated and dynamic financial market presents a persistent challenge. Fully capturing various market variables, inc…

cs.CL2025

Unleashing the potential of prompt engineering for large language models

Banghao Chen, Zhaofeng Zhang, Nicolas Langrené +1

This comprehensive review delves into the pivotal role of prompt engineering in unleashing the capabilities of Large Language Models (LLMs). The development of Artificial Intellige…