9 citations · 11 across the 9 of their papers we have counts for
9 papers
CATCH: Complementary Adaptive Token-level Contrastive Decoding to Mitigate Hallucinations in LVLMs
Zhehan Kan, Ce Zhang, Zihan Liao +7
Large Vision-Language Model (LVLM) systems have demonstrated impressive vision-language reasoning capabilities but suffer from pervasive and severe hallucination issues, posing sig…
UV Gaussians: Joint Learning of Mesh Deformation and Gaussian Textures for Human Avatar Modeling
Yujiao Jiang, Qingmin Liao, Xiaoyu Li +5
Reconstructing photo-realistic drivable human avatars from multi-view image sequences has been a popular and challenging topic in the field of computer vision and graphics. While e…
Localization matters too: How localization error affects UAV flight
Suquan Zhang, Yuanfan Xu, Shu'ang Yu +3
The maximum safe flight speed of a Unmanned Aerial Vehicle (UAV) is an important indicator for measuring its efficiency in completing various tasks. This indicator is influenced by…
DiffVein: A Unified Diffusion Network for Finger Vein Segmentation and Authentication
Yanjun Liu, Wenming Yang, Qingmin Liao
Finger vein authentication, recognized for its high security and specificity, has become a focal point in biometric research. Traditional methods predominantly concentrate on vein…
LLM-Powered Hierarchical Language Agent for Real-time Human-AI Coordination
Jijia Liu, Chao Yu, Jiaxuan Gao +4
AI agents powered by Large Language Models (LLMs) have made significant advances, enabling them to assist humans in diverse complex tasks and leading to a revolution in human-AI co…
DSR-Diff: Depth Map Super-Resolution with Diffusion Model
Yuan Shi, Bin Xia, Rui Zhu +2
Color-guided depth map super-resolution (CDSR) improve the spatial resolution of a low-quality depth map with the corresponding high-quality color map, benefiting various applicati…