activity
20242026
most citedTowards Explainable Fake Image Detection with Multi-Modal Large Language Models

2 citations · 2 across the 10 of their papers we have counts for

collaborators
Showing 2024Show all

5 papers · 1 filter

cs.CV2024

DomainGallery: Few-shot Domain-driven Image Generation by Attribute-centric Finetuning

Yuxuan Duan, Yan Hong, Bo Zhang +6

The recent progress in text-to-image models pretrained on large-scale datasets has enabled us to generate various images as long as we provide a text prompt describing what we want…

cs.CV2024

DeMamba: AI-Generated Video Detection on Million-Scale GenVideo Benchmark

Haoxing Chen, Yan Hong, Zizheng Huang +8

Recently, video generation techniques have advanced rapidly. Given the popularity of video content on social media platforms, these models intensify concerns about the spread of fa…

cs.CV2024

Conditional Prototype Rectification Prompt Learning

Haoxing Chen, Yaohui Li, Zizheng Huang +6

Pre-trained large-scale vision-language models (VLMs) have acquired profound understanding of general visual concepts. Recent advancements in efficient transfer learning (ETL) have…

cs.CV2024

Supervised Contrastive Learning for Snapshot Spectral Imaging Face Anti-Spoofing

Chuanbiao Song, Yan Hong, Jun Lan +3

This study reveals a cutting-edge re-balanced contrastive learning strategy aimed at strengthening face anti-spoofing capabilities within facial recognition systems, with a focus o…

cs.CV2024

Boosting Audio-visual Zero-shot Learning with Large Language Models

Haoxing Chen, Yaohui Li, Yan Hong +6

Audio-visual zero-shot learning aims to recognize unseen classes based on paired audio-visual sequences. Recent methods mainly focus on learning multi-modal features aligned with c…