3 citations · 3 across the 2 of their papers we have counts for
3 papers
cs.LG2026
PithTrain: A Compact and Agent-Native MoE Training System
Ruihang Lai, Hao Kang, Haozhan Tang +6
Mixture-of-Experts (MoE) has become the dominant architecture for frontier language models. To meet this demand, production frameworks have built optimized MoE training stacks over…
cs.CL2026★ 3 cited
Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization
Tzu-Quan Lin, Wei-Ping Huang, Hao Tang +1
Fine-tuning speech representation models can enhance performance on specific tasks but often compromises their cross-task generalization ability. This degradation is often caused b…
cs.CL2025
Identifying Speaker Information in Feed-Forward Layers of Self-Supervised Speech Transformers
Tzu-Quan Lin, Hsi-Chun Cheng, Hung-yi Lee +1
In recent years, the impact of self-supervised speech Transformers has extended to speaker-related applications. However, little research has explored how these models encode speak…