3 papers
cs.LG2026
Gecko: Fast Private Inference via Secure Public Encoder Offloading
Cheng'an Wei, Kai Chen, Yue Zhao +2
Private inference protects both user inputs and server models during neural network inference, but existing solutions remain too slow for practical deployment. This motivates recen…
cs.AI2024
Hidden in Plain Sight: Exploring Chat History Tampering in Interactive Language Models
Cheng'an Wei, Yue Zhao, Yujia Gong +3
Large Language Models (LLMs) such as ChatGPT and Llama have become prevalent in real-world applications, exhibiting impressive text generation performance. LLMs are fundamentally d…
cs.CL2023
LLM Factoscope: Uncovering LLMs' Factual Discernment through Inner States Analysis
Jinwen He, Yujia Gong, Kai Chen +3
Large Language Models (LLMs) have revolutionized various domains with extensive knowledge and creative capabilities. However, a critical issue with LLMs is their tendency to produc…