5 papers
Fast Model Selection and Stable Optimization for Softmax-Gated Multinomial-Logistic Mixture of Experts Models
TrungKhang Tran, TrungTin Nguyen, Md Abul Bashar +3
Mixture-of-Experts (MoE) architectures combine specialized predictors through a learned gate and are effective across regression and classification, but for classification with sof…
Revisiting Incremental Stochastic Majorization-Minimization Algorithms with Applications to Mixture of Experts
TrungKhang Tran, TrungTin Nguyen, Gersende Fort +5
Processing high-volume, streaming data is increasingly common in modern statistics and machine learning, where batch-mode algorithms are often impractical because they require repe…
LGCA: Enhancing Semantic Representation via Progressive Expansion
Thanh Hieu Cao, Trung Khang Tran, Gia Thinh Pham +2
Recent advancements in large-scale pretraining in natural language processing have enabled pretrained vision-language models such as CLIP to effectively align images and text, sign…
Communication Bounds for the Distributed Experts Problem
Zhihao Jia, Qi Pang, Trung Tran +3
In this work, we study the experts problem in the distributed setting where an expert's cost needs to be aggregated across multiple servers. Our study considers various communicati…
Matrix Completion in Group Testing: Bounds and Simulations
Trung-Khang Tran, Thach V. Bui
The goal of group testing is to identify a small number of defective items within a large population. In the non-adaptive setting, tests are designed in advance and represented by…