Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
A Primal-Dual Framework for Transformers and Neural Networks
Tan M. Nguyen, Tam Nguyen, Nhat Ho +3
Self-attention is key to the remarkable success of transformers in sequence modeling tasks including many applications in natural language processing and computer vision. Like neur…
cs.LG2024
Operator Splitting for Learning to Predict Equilibria in Convex Games
Daniel McKenzie, Howard Heaton, Qiuwei Li +3
Systems of competing agents can often be modeled as games. Assuming rationality, the most likely outcomes are given by an equilibrium (e.g. a Nash equilibrium). In many practical s…