1 citations · 1 across the 4 of their papers we have counts for
4 papers
Online DPO: Online Direct Preference Optimization with Fast-Slow Chasing
Biqing Qi, Pengfei Li, Fangyuan Li +3
Direct Preference Optimization (DPO) improves the alignment of large language models (LLMs) with human values by training directly on human preference datasets, eliminating the nee…
SMR: State Memory Replay for Long Sequence Modeling
Biqing Qi, Junqi Gao, Kaiyan Zhang +4
Despite the promising performance of state space models (SSMs) in long sequence modeling, limitations still exist. Advanced SSMs like S5 and S6 (Mamba) in addressing non-uniform sa…
Interactive Continual Learning: Fast and Slow Thinking
Biqing Qi, Xingquan Chen, Junqi Gao +4
Advanced life forms, sustained by the synergistic interaction of neural cognitive mechanisms, continually acquire and transfer knowledge throughout their lifespan. In contrast, con…
Contrastive Augmented Graph2Graph Memory Interaction for Few Shot Continual Learning
Biqing Qi, Junqi Gao, Xingquan Chen +4
Few-Shot Class-Incremental Learning (FSCIL) has gained considerable attention in recent years for its pivotal role in addressing continuously arriving classes. However, it encounte…