2 papers
cs.LG2025
Optimizing MoE Routers: Design, Implementation, and Evaluation in Transformer Models
Daniel Fidel Harvey, George Weale, Berk Yilmaz
Mixture of Experts (MoE) architectures increase large language model scalability, yet their performance depends on the router module that moves tokens to specialized experts. Bad r…
cs.CV2025
Efficient Transformations in Deep Learning Convolutional Neural Networks
Berk Yilmaz, Daniel Fidel Harvey, Prajit Dhuri
This study investigates the integration of signal processing transformations -- Fast Fourier Transform (FFT), Walsh-Hadamard Transform (WHT), and Discrete Cosine Transform (DCT) --…