activity
20192026
most citedTowards Transfer Learning for End-to-End Speech Synthesis from Deep Pre-Trained Language Models

26 citations · 42 across the 9 of their papers we have counts for

collaborators
Showing cs.CLShow all

7 papers · 1 filter

cs.CL2026

Beyond Single-Shot: Multi-step Tool Retrieval via Query Planning

Wei Fang, James Glass

LLM agents operating over massive, dynamic tool libraries rely on effective retrieval, yet standard single-shot dense retrievers struggle with complex requests. These failures prim…

cs.CL2026★ 1 cited

MedDialogRubrics: A Comprehensive Benchmark and Evaluation Framework for Multi-turn Medical Consultations in Large Language Models

Lecheng Gong, Weimin Fang, Ting Yang +9

Medical conversational AI (AI) plays a pivotal role in the development of safer and more effective medical dialogue systems. However, existing benchmarks and evaluation frameworks…

cs.CL2025

Med-RewardBench: Benchmarking Reward Models and Judges for Medical Multimodal Large Language Models

Meidan Ding, Jipeng Zhang, Wenxuan Wang +6

Multimodal large language models (MLLMs) hold significant potential in medical applications, including disease diagnosis and clinical decision-making. However, these tasks require…

cs.CL2023★ 2 cited

Expand, Rerank, and Retrieve: Query Reranking for Open-Domain Question Answering

Yung-Sung Chuang, Wei Fang, Shang-Wen Li +2

We propose EAR, a query Expansion And Reranking approach for improving passage retrieval, with the application to open-domain question answering. EAR first applies a query expansio…

cs.CL2023★ 8 cited

Interpretable Unified Language Checking

Tianhua Zhang, Hongyin Luo, Yung-Sung Chuang +7

Despite recent concerns about undesirable behaviors generated by large language models (LLMs), including non-factual, biased, and hateful language, we find LLMs are inherent multi-…

cs.CL2019★ 26 cited

Towards Transfer Learning for End-to-End Speech Synthesis from Deep Pre-Trained Language Models

Wei Fang, Yu-An Chung, James Glass

Modern text-to-speech (TTS) systems are able to generate audio that sounds almost as natural as human speech. However, the bar of developing high-quality TTS systems remains high s…