10 citations · 12 across the 4 of their papers we have counts for
4 papers
WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning
Yuanhan Zhang, Kaichen Zhang, Bo Li +4
Multimodal information, together with our knowledge, help us to understand the complex and dynamic world. Large language models (LLM) and large multimodal models (LMM), however, st…
Robust Face Anti-Spoofing with Dual Probabilistic Modeling
Yuanhan Zhang, Yichao Wu, Zhenfei Yin +2
The field of face anti-spoofing (FAS) has witnessed great progress with the surge of deep learning. Due to its data-driven nature, existing FAS methods are sensitive to the noise i…
RealNet: Combining Optimized Object Detection with Information Fusion Depth Estimation Co-Design Method on IoT
Zhuohao Li, Fandi Gou, Qixin De +3
Depth Estimation and Object Detection Recognition play an important role in autonomous driving technology under the guidance of deep learning artificial intelligence. We propose a…
CelebA-Spoof Challenge 2020 on Face Anti-Spoofing: Methods and Results
Yuanhan Zhang, Zhenfei Yin, Jing Shao +22
As facial interaction systems are prevalently deployed, security and reliability of these systems become a critical issue, with substantial research efforts devoted. Among them, fa…