From the 1 of 4 linked papers with an AI index.
4 papers
Equitable System-Prompt Selection via Constrained Mixed-Strategy GroupDRO
Mengyu Xu, Qiaoxin Yang, Zhihan Liu +4
Large language models are increasingly used for information seeking, yet semantically equivalent questions phrased in different ways can receive answers of considerably different q…
Rethinking LLM-Judged Helpfulness as a Pedagogy Signal: A Pre-Registered Audit Across Tutor Models
Shuyi Fan, Boyuan Deng, Mengyu Xu +4
The paper audits whether a general helpfulness rubric can reliably distinguish between answer‑giving and pedagogical guidance in large language model tutoring, finding that helpful…
LoopFM: Learning frOm HistOrical RePresentations of Foundation Model for Recommendation
Shali Jiang, Hua Zheng, Boyang Liu +40
Knowledge distillation (KD) transfers a single scalar prediction from a large foundation model (FM) to compact vertical models (VMs), suffering from diminishing transfer ratio -- t…
MIRA: A Bilingual Benchmark for Medical Information Response Audit
Mengyu Xu, Qiaoxin Yang, Qianqian Wang +3
Large language models (LLMs) are increasingly used to provide public-facing health information, yet existing safety evaluations overlook whether responses preserve comparable medic…