5 papers
LoCA: Forward-Only LLM Tuning after One-Shot Calibration with Local Credit Assignment
Linhan Xia, Rui Liu, Zhaofeng Zhang +3
Parameter-efficient post-training reduces the number of trainable parameters, but still requires repeated end-to-end backpropagation through the frozen backbone. Every adaptation s…
Evolving in the Agent Jungle via History-Informed Opponent Awareness
Zhaofeng Zhang, Linhan Xia, Rui Liu +3
Learning to adapt strategies through interaction is a key step toward more general and autonomous LLM agents. Existing approaches typically achieve behavioral adaptation by revisin…
Adaptive Dueling Double Deep Q-networks in Uniswap V3 Replication and Extension with Mamba
Zhaofeng Zhang
The report goes through the main steps of replicating and improving the article "Adaptive Liquidity Provision in Uniswap V3 with Deep Reinforcement Learning." The replication part…
Quantformer: from attention to profit with a quantitative transformer trading strategy
Zhaofeng Zhang, Banghao Chen, Shengxin Zhu +1
In traditional quantitative trading practice, navigating the complicated and dynamic financial market presents a persistent challenge. Fully capturing various market variables, inc…
Unleashing the potential of prompt engineering for large language models
Banghao Chen, Zhaofeng Zhang, Nicolas Langrené +1
This comprehensive review delves into the pivotal role of prompt engineering in unleashing the capabilities of Large Language Models (LLMs). The development of Artificial Intellige…