activity
20162026
most citedWeakly Aligned Feature Fusion for Multimodal Object Detection

97 citations · 411 across the 56 of their papers we have counts for

collaborators
Showing 2024Show all

13 papers · 1 filter

cs.CV2024

MVBoost: Boost 3D Reconstruction with Multi-View Refinement

Xiangyu Liu, Xiaomei Zhang, Zhiyuan Ma +2

Recent advancements in 3D object reconstruction have been remarkable, yet most current 3D models rely heavily on existing 3D datasets. The scarcity of diverse 3D datasets results i…

cs.CV2024

Spoof Trace Discovery for Deep Learning Based Explainable Face Anti-Spoofing

Haoyuan Zhang, Xiangyu Zhu, Li Gao +4

With the rapid growth usage of face recognition in people's daily life, face anti-spoofing becomes increasingly important to avoid malicious attacks. Recent face anti-spoofing mode…

cs.CV2024★ 17 cited

Second FRCSyn-onGoing: Winning Solutions and Post-Challenge Analysis to Improve Face Recognition with Synthetic Data

Ivan DeAndres-Tame, Ruben Tolosana, Pietro Melzi +56

Synthetic data is gaining increasing popularity for face recognition technologies, mainly due to the privacy concerns and challenges associated with obtaining real data, including…

cs.CV2024

Revisiting Marr in Face: The Building of 2D--2.5D--3D Representations in Deep Neural Networks

Xiangyu Zhu, Chang Yu, Jiankuo Zhao +3

David Marr's seminal theory of vision proposes that the human visual system operates through a sequence of three stages, known as the 2D sketch, the 2.5D sketch, and the 3D model.…

cs.CV2024

Enhancing Instruction-Following Capability of Visual-Language Models by Reducing Image Redundancy

Te Yang, Jian Jia, Xiangyu Zhu +9

Large Language Models (LLMs) have strong instruction-following capability to interpret and execute tasks as directed by human commands. Multimodal Large Language Models (MLLMs) hav…

cs.CV2024

S2TD-Face: Reconstruct a Detailed 3D Face with Controllable Texture from a Single Sketch

Zidu Wang, Xiangyu Zhu, Jiang Yu +2

3D textured face reconstruction from sketches applicable in many scenarios such as animation, 3D avatars, artistic design, missing people search, etc., is a highly promising but un…