Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Unified CNNs and transformers underlying learning mechanism reveals multi-head attention modus vivendi
Ella Koresh, Ronit D. Gross, Yuval Meir +3
Convolutional neural networks (CNNs) evaluate short-range correlations in input images which progress along the layers, whereas vision transformer (ViT) architectures evaluate long…
cs.LG2025
Advanced deep architecture pruning using single filter performance
Yarden Tzach, Yuval Meir, Ronit D. Gross +3
Pruning the parameters and structure of neural networks reduces the computational complexity, energy consumption, and latency during inference. Recently, a novel underlying mechani…