3 papers
cs.LG2026
Multimodal Unlearning Across Vision, Language, Video, and Audio: Survey of Methods, Datasets, and Benchmarks
Nobin Sarwar, Shubhashis Roy Dipta, Zheyuan Liu +1
With the growing adoption of VLMs, DMs, LLMs, and AFMs, these multimodal foundation models can inadvertently encode sensitive, copyrighted, biased, or unsafe cross-modal associatio…
cs.CV2026
Exploiting the Final Component of Generator Architectures for AI-Generated Image Detection
Yanzhu Liu, Xiao Liu, Yuexuan Wang +1
With the rapid proliferation of powerful image generators, accurate detection of AI-generated images has become essential for maintaining a trustworthy online environment. However,…
cs.LG2025
Variance-Reduction Guidance: Sampling Trajectory Optimization for Diffusion Models
Shifeng Xu, Yanzhu Liu, Adams Wai-Kin Kong
Diffusion models have become emerging generative models. Their sampling process involves multiple steps, and in each step the models predict the noise from a noisy sample. When the…