Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
SimulatorArena: Are User Simulators Reliable Proxies for Multi-Turn Evaluation of AI Assistants?
Yao Dou, Michel Galley, Baolin Peng +6
Large language models (LLMs) are increasingly used in interactive applications, and human evaluation remains the gold standard for assessing their performance in multi-turn convers…
cs.CL2024
SoftTiger: A Clinical Foundation Model for Healthcare Workflows
Ye Chen, Igor Couto, Wei Cai +2
We introduce SoftTiger, a clinical large language model (CLaM) designed as a foundation model for healthcare workflows. The narrative and unstructured nature of clinical notes is a…
cs.CL2023
Joint Repetition Suppression and Content Moderation of Large Language Models
Minghui Zhang, Alex Sokolov, Weixin Cai +1
Natural language generation (NLG) is one of the most impactful fields in NLP, and recent years have witnessed its evolution brought about by large language models (LLMs). As the ke…