activity
20242026
collaborators

16 papers

cs.LG2026

AsFT: Anchoring Safety During LLM Fine-Tuning Within Narrow Safety Basin

Shuo Yang, Qihui Zhang, Yuyang Liu +7

Fine-tuning large language models (LLMs) improves performance but introduces critical safety vulnerabilities: even minimal harmful data can severely compromise safety measures. We…

cs.CL2026

RankLLM: Weighted Ranking of LLMs by Quantifying Question Difficulty

Ziqian Zhang, Xingjian Hu, Yue Huang +8

Benchmarks establish a standardized evaluation framework to systematically assess the performance of large language models (LLMs), facilitating objective comparisons and driving ad…

cs.CY2026

On the Trustworthiness of Generative Foundation Models: Guideline, Assessment, and Perspective

Yue Huang, Chujie Gao, Siyuan Wu +63

Generative Foundation Models (GenFMs) have emerged as transformative tools. However, their widespread adoption raises critical concerns regarding trustworthiness across dimensions.…

cs.CL2026

Dep-Search: Learning Dependency-Aware Reasoning Traces with Persistent Memory

Yanming Liu, Xinyue Peng, Zixuan Yan +7

Large Language Models (LLMs) have demonstrated remarkable capabilities in complex reasoning tasks, particularly when augmented with search mechanisms that enable systematic explora…

cs.AI2026

Digital Twin AI: Opportunities and Challenges from Large Language Models to World Models

Rong Zhou, Dongping Chen, Zihan Jia +24

Digital twins, as precise digital representations of physical systems, have evolved from passive simulation tools into intelligent and autonomous entities through the integration o…

cs.CL2025

DeID-GPT: Zero-shot Medical Text De-Identification by GPT-4

Zhengliang Liu, Yue Huang, Xiaowei Yu +15

The digitization of healthcare has facilitated the sharing and re-using of medical data but has also raised concerns about confidentiality and privacy. HIPAA (Health Insurance Port…