2 papers
cs.CV2026
MedTri: A Platform for Structured Medical Report Normalization to Enhance Vision-Language Pretraining
Yuetan Chu, Xinhua Ma, Xinran Jin +2
Medical vision-language pretraining increasingly relies on medical reports as large-scale supervisory signals; however, raw reports often exhibit substantial stylistic heterogeneit…
cs.LG2025
MegaScale-MoE: Large-Scale Communication-Efficient Training of Mixture-of-Experts Models in Production
Chao Jin, Ziheng Jiang, Zhihao Bai +16
We present MegaScale-MoE, a production system tailored for the efficient training of large-scale mixture-of-experts (MoE) models. MoE emerges as a promising architecture to scale l…