Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Adaptive Sparse Softmax: An Effective and Efficient Softmax Variant
Qi Lv, Lei Geng, Ziqiang Cao +4
Softmax with the cross entropy loss is the standard configuration for current neural classification models. The gold score for a target class is supposed to be 1, but it is never r…
cs.LG2025
Decision Mamba: A Multi-Grained State Space Model with Self-Evolution Regularization for Offline RL
Qi Lv, Xiang Deng, Gongwei Chen +2
While the conditional sequence modeling with the transformer architecture has demonstrated its effectiveness in dealing with offline reinforcement learning (RL) tasks, it is strugg…