activity
20222024
most citedAlleviating Structural Distribution Shift in Graph Anomaly Detection

66 citations · 268 across the 18 of their papers we have counts for

collaborators
Showing cs.CVShow all

6 papers · 1 filter

cs.CV20242 cited

GlanceVAD: Exploring Glance Supervision for Label-efficient Video Anomaly Detection

Huaxin Zhang, Xiang Wang, Xiaohao Xu +6

In recent years, video anomaly detection has been extensively investigated in both unsupervised and weakly supervised settings to alleviate costly temporal labeling. Despite signif…

cs.CV202322 cited

I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models

Shiwei Zhang, Jiayu Wang, Yingya Zhang +6

Video synthesis has recently made remarkable strides benefiting from the rapid development of diffusion models. However, it still encounters challenges in terms of semantic accurac…

cs.CV202347 cited

ModelScope Text-to-Video Technical Report

Jiuniu Wang, Hangjie Yuan, Dayou Chen +3

This paper introduces ModelScopeT2V, a text-to-video synthesis model that evolves from a text-to-image synthesis model (i.e., Stable Diffusion). ModelScopeT2V incorporates spatio-t…

cs.CV202322 cited

Redundancy-aware Transformer for Video Question Answering

Yicong Li, Xun Yang, An Zhang +3

This paper identifies two kinds of redundancy in the current VideoQA paradigm. Specifically, the current video encoders tend to holistically embed all video clues at different gran…

cs.CV20233 cited

MoLo: Motion-augmented Long-short Contrastive Learning for Few-shot Action Recognition

Xiang Wang, Shiwei Zhang, Zhiwu Qing +4

Current state-of-the-art approaches for few-shot action recognition achieve promising performance by conducting frame-level matching on learned visual features. However, they gener…

cs.CV20232 cited

HyRSM++: Hybrid Relation Guided Temporal Set Matching for Few-shot Action Recognition

Xiang Wang, Shiwei Zhang, Zhiwu Qing +4

Recent attempts mainly focus on learning deep representations for each video individually under the episodic meta-learning regime and then performing temporal alignment to match qu…