1 citations · 1 across the 2 of their papers we have counts for
5 papers
Neural Gate: Mitigating Privacy Risks in LVLMs via Neuron-Level Gradient Gating
Xiangkui Cao, Jie Zhang, Meina Kan +2
Large Vision-Language Models (LVLMs) have shown remarkable potential across a wide array of vision-language tasks, leading to their adoption in critical domains such as finance and…
Plan-R1: Safe and Feasible Trajectory Planning as Language Modeling
Xiaolong Tang, Meina Kan, Shiguang Shan +1
Safe and feasible trajectory planning is critical for real-world autonomous driving systems. However, existing learning-based planners rely heavily on expert demonstrations, which…
OSI: One-step Inversion Excels in Extracting Diffusion Watermarks
Yuwei Chen, Zhenliang He, Jia Tang +2
Watermarking is an important mechanism for provenance and copyright protection of diffusion-generated images. Training-free methods, exemplified by Gaussian Shading, embed watermar…
JoPano: Unified Panorama Generation via Joint Modeling
Wancheng Feng, Chen An, Zhenliang He +3
Panorama generation has recently attracted growing interest in the research community, with two core tasks, text-to-panorama and view-to-panorama generation. However, existing meth…
Jodi: Unification of Visual Generation and Understanding via Joint Modeling
Yifeng Xu, Zhenliang He, Meina Kan +2
Visual generation and understanding are two deeply interconnected aspects of human intelligence, yet they have been traditionally treated as separate tasks in machine learning. In…