8 papers
FLaG: Fine-Grained Latent Grouping for Hallucination Detection
Wentao Ye, Liyao Li, Zhiqing Xiao +6
Hallucinations in large language models (LLMs) arise from heterogeneous failure mechanisms, making reliable detection difficult for any single global uncertainty score. In this wor…
KMLP: A Scalable Hybrid Architecture for Web-Scale Tabular Data Modeling
Mingming Zhang, Pengfei Shi, Zhiqing Xiao +8
Predictive modeling on web-scale tabular data with billions of instances and hundreds of heterogeneous numerical features faces significant scalability challenges. These features e…
An Invariant Latent Space Perspective on Language Model Inversion
Wentao Ye, Jiaqi Hu, Haobo Wang +7
Language model inversion (LMI), i.e., recovering hidden prompts from outputs, emerges as a concrete threat to user privacy and system security. We recast LMI as reusing the LLM's o…
SPA++: Generalized Graph Spectral Alignment for Versatile Domain Adaptation
Zhiqing Xiao, Haobo Wang, Xu Lu +3
Domain Adaptation (DA) aims to transfer knowledge from a labeled source domain to an unlabeled or sparsely labeled target domain under domain shifts. Most prior works focus on capt…
ALPS: Attention Localization and Pruning Strategy for Efficient Alignment of Large Language Models
Hao Chen, Haoze Li, Zhiqing Xiao +6
Aligning general-purpose large language models (LLMs) to downstream tasks often incurs significant training adjustment costs. Prior research has explored various avenues to enhance…
Jailbreaking Prompt Attack: A Controllable Adversarial Attack against Diffusion Models
Jiachen Ma, Yijiang Li, Zhiqing Xiao +4
Text-to-image (T2I) models can be maliciously used to generate harmful content such as sexually explicit, unfaithful, and misleading or Not-Safe-for-Work (NSFW) images. Previous at…