6 papers
JND-Guided Neural Watermarking with Spatial Transformer Decoding for Screen-Capture Robustness
Jiayi Qin, Jingwei Li, Chuan Wu
Screen-shooting robust watermarking aims to imperceptibly embed extractable information into host images such that the watermark survives the complex distortion pipeline of screen…
Efficient Image Super-Resolution with Multi-Scale Spatial Adaptive Attention Networks
Sushi Rao, Jingwei Li
This paper introduces a lightweight image super-resolution (SR) network, termed the Multi-scale Spatial Adaptive Attention Network (MSAAN), to address the common dilemma between hi…
Detecting Deepfakes with Multivariate Soft Blending and CLIP-based Image-Text Alignment
Jingwei Li, Jiaxin Tong, Pengfei Wu
The proliferation of highly realistic facial forgeries necessitates robust detection methods. However, existing approaches often suffer from limited accuracy and poor generalizatio…
The Diffusion Duet: Harmonizing Dual Channels with Wavelet Suppression for Image Separation
Jingwei Li, Wei Pu
Blind image separation (BIS) refers to the inverse problem of simultaneously estimating and restoring multiple independent source images from a single observation image under condi…
Kimi K2.5: Visual Agentic Intelligence
Kimi Team, Tongtong Bai, Yifan Bai +339
We introduce Kimi K2.5, an open-source multimodal agentic model designed to advance general agentic intelligence. K2.5 emphasizes the joint optimization of text and vision so that…
MutualNeRF: Improve the Performance of NeRF under Limited Samples with Mutual Information Theory
Zifan Wang, Jingwei Li, Yitang Li +1
This paper introduces MutualNeRF, a framework enhancing Neural Radiance Field (NeRF) performance under limited samples using Mutual Information Theory. While NeRF excels in 3D scen…