2 papers
cs.LG2025
Mixture of Raytraced Experts
Andrea Perin, Giacomo Lagomarsini, Claudio Gallicchio +1
We introduce a Mixture of Raytraced Experts, a stacked Mixture of Experts (MoE) architecture which can dynamically select sequences of experts, producing computational graphs of va…
cs.LG2024
On the Ability of Deep Networks to Learn Symmetries from Data: A Neural Kernel Theory
Andrea Perin, Stephane Deny
Symmetries (transformations by group actions) are present in many datasets, and leveraging them holds considerable promise for improving predictions in machine learning. In this wo…