From the 1 of 10 linked papers with an AI index.
10 papers
The Hyperspherical Geometry of CLIP Latent Space: A Semantic Mixture Model
Zijie Yu, Gaowen Liu, Ramana Rao Kompella +2
The paper introduces a probabilistic model for CLIP embeddings using mixtures of von Mises-Fisher distributions on the unit hypersphere, improving density estimation and detection…
Safeguarding Text-to-Image Generation via Inference-Time Prompt-Noise Optimization
Jiangweizhi Peng, Zhiwei Tang, Gaowen Liu +2
Text-to-Image (T2I) diffusion models are widely recognized for their ability to generate high-quality and diverse images based on text prompts. However, despite recent advances, th…
MetaSeal: Defending Against Image Attribution Forgery Through Content-Dependent Cryptographic Watermarks
Tong Zhou, Ruyi Ding, Gaowen Liu +5
The rapid growth of digital and AI-generated images has amplified the need for secure and verifiable methods of image attribution. While digital watermarking offers more robust pro…
Do LLMs Recognize Your Latent Preferences? A Benchmark for Latent Information Discovery in Personalized Interaction
Ioannis Tsaknakis, Bingqing Song, Shuyu Gan +5
Large Language Models (LLMs) excel at producing broadly relevant text, but this generality becomes a limitation when user-specific preferences are required, such as recommending re…
ConQuER: Modular Architectures for Control and Bias Mitigation in IQP Quantum Generative Models
Xiaocheng Zou, Shijin Duan, Charles Fleming +4
Quantum generative models based on instantaneous quantum polynomial (IQP) circuits show great promise in learning complex distributions while maintaining classical trainability. Ho…
A Neurosymbolic Agent System for Compositional Visual Reasoning
Yichang Xu, Gaowen Liu, Ramana Rao Kompella +5
The advancement in large language models (LLMs) and large vision models has fueled the rapid progress in multi-modal vision-language reasoning capabilities. However, existing visio…