collaborators

6 papers

cs.CL2025

47B Mixture-of-Experts Beats 671B Dense Models on Chinese Medical Examinations

Chiung-Yi Tseng, Danyang Zhang, Tianyang Wang +8

The rapid advancement of large language models(LLMs) has prompted significant interest in their potential applications in medical domains. This paper presents a comprehensive bench…

cs.CL2025

Affective Multimodal Agents with Proactive Knowledge Grounding for Emotionally Aligned Marketing Dialogue

Lin Yu, Xiaofei Han, Yifei Kang +4

Recent advances in large language models (LLMs) have enabled fluent dialogue systems, but most remain reactive and struggle in emotionally rich, goal-oriented settings such as mark…

cs.LG2025

Is GPT-OSS All You Need? Benchmarking Large Language Models for Financial Intelligence and the Surprising Efficiency Paradox

Ziqian Bi, Danyang Zhang, Junhao Song +1

The rapid adoption of large language models in financial services necessitates rigorous evaluation frameworks to assess their performance, efficiency, and practical applicability.…

cs.CL2025

StreetMath: Study of LLMs' Approximation Behaviors

Chiung-Yi Tseng, Somshubhra Roy, Maisha Thasin +2

There is a substantial body of literature examining the mathematical reasoning capabilities of large language models (LLMs), particularly their performance on precise arithmetic op…

cs.CL2025

Is GPT-OSS Good? A Comprehensive Evaluation of OpenAI's Latest Open Source Models

Ziqian Bi, Keyu Chen, Chiung-Yi Tseng +9

In August 2025, OpenAI released GPT-OSS models, its first open weight large language models since GPT-2 in 2019, comprising two mixture of experts architectures with 120B and 20B p…

cs.LG2025

Mixture of Experts in Large Language Models

Danyang Zhang, Junhao Song, Ziqian Bi +5

This paper presents a comprehensive review of the Mixture-of-Experts (MoE) architecture in large language models, highlighting its ability to significantly enhance model performanc…