1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.LG2025
Maximum Score Routing For Mixture-of-Experts
Bowen Dong, Yilong Fan, Yutao Sun +4
Routing networks in sparsely activated mixture-of-experts (MoE) dynamically allocate input tokens to top-k experts through differentiable sparse transformations, enabling scalable…
cs.AR2024★ 1 cited
DEFA: Efficient Deformable Attention Acceleration via Pruning-Assisted Grid-Sampling and Multi-Scale Parallel Processing
Yansong Xu, Dongxu Lyu, Zhenyu Li +6
Multi-scale deformable attention (MSDeformAttn) has emerged as a key mechanism in various vision tasks, demonstrating explicit superiority attributed to multi-scale grid-sampling.…
cs.CV2022
LiteDepth: Digging into Fast and Accurate Depth Estimation on Mobile Devices
Zhenyu Li, Zehui Chen, Jialei Xu +2
Monocular depth estimation is an essential task in the computer vision community. While tremendous successful methods have obtained excellent results, most of them are computationa…