2 papers
cs.LG2026
Fast MoE Inference via Predictive Prefetching and Expert Replication
Ankit Jyothish, Ali Jannesari, Aishwarya Sarkar +1
The Mixture of Experts (MoE) architecture has become a fundamental building block in state-of-the-art large language models (LLMs), improving domain-specific expertise in LLMs and…
cs.LG2025
Leveraging Manifold Embeddings for Enhanced Graph Transformer Representations and Learning
Ankit Jyothish, Ali Jannesari
Graph transformers typically embed every node in a single Euclidean space, blurring heterogeneous topologies. We prepend a lightweight Riemannian mixture-of-experts layer that rout…