3 papers
cs.CV2026
Unison: Benchmarking Unified Multimodal Models via Synergistic Understanding and Generation
Jinyu Liu, Xincheng Shuai, Henghui Ding +1
Unified multimodal models capable of both understanding and generation have achieved remarkable strides. However, despite their unified designs, existing evaluations typically asse…
cs.CV2026
Real-Time Glottis Detection Framework via Spatial-decoupled Feature Learning for Nasal Transnasal Intubation
Jinyu Liu, Gaoyang Zhang, Yang Zhou +3
Nasotracheal intubation (NTI) is a vital procedure in emergency airway management, where rapid and accurate glottis detection is essential to ensure patient safety. However, existi…
cs.CV2025
Ref-SAM3D: Bridging SAM3D with Text for Reference 3D Reconstruction
Yun Zhou, Yaoting Wang, Guangquan Jie +2
SAM3D has garnered widespread attention for its strong 3D object reconstruction capabilities. However, a key limitation remains: SAM3D cannot reconstruct specific objects referred…