3 papers
cs.CV2025
Aligning First, Then Fusing: A Novel Weakly Supervised Multimodal Violence Detection Method
Wenping Jin, Li Zhu, Jing Sun
Weakly supervised violence detection refers to the technique of training models to identify violent segments in videos using only video-level labels. Among these approaches, multim…
cs.CV2024
DMSD-CDFSAR: Distillation from Mixed-Source Domain for Cross-Domain Few-shot Action Recognition
Fei Guo, YiKang Wang, Han Qi +2
Few-shot action recognition is an emerging field in computer vision, primarily focused on meta-learning within the same domain. However, challenges arise in real-world scenario dep…
cs.CV2023
Task-Specific Alignment and Multiple Level Transformer for Few-Shot Action Recognition
Fei Guo, Li Zhu, YiWang Wang +1
In the research field of few-shot learning, the main difference between image-based and video-based is the additional temporal dimension. In recent years, some works have used the…