output
20062026
most citedMotion Planning for Autonomous Driving: The State of the Art and Future Perspectives

583 citations

172 papers

cs.CV2026

Audio-Visual Segmentation via Depth-Guided Collaborative Modeling

Zhaojin Fu, Yuyang Hong, Qi Yang +4

Audio-Visual Segmentation (AVS) is a fundamental task in multimodal perception that performs pixel-level segmentation of sounding objects in videos by leveraging both visual and au…

cs.AI2026

SymDiag: Explainable Diagnosis for LLM Reasoning via Neuro-Symbolic Verification

Wenyao Cui, Huaping Zhang, Yongyi Huang +6

Large language models (LLMs) increasingly serve as data-driven reasoners, yet their chains-of-thought (CoT) can be unfaithful even when final answers are correct. Most existing ``v…

cs.CV2026

Geometry-Aware Camera Localization for Bronchoscopy

Lumin Chen, Qingyao Tian, Jinpeng Li +5

Camera localization in bronchoscopy remains a challenging problem due to stringent accuracy requirements, real-time constraints, and limited training data. Compared to natural scen…

cs.CV2026

Nexus: Native Mesh Generation with Diffusion

Hanxiao Wang, Ying-Tian Liu, Yuan-Chen Guo +5

Generating high-quality triangle meshes is essential for film, gaming, and interactive 3D applications. Mainstream methods rely on mesh serialization and autoregressive processes,…

cs.CV2026

Disentanglement-Based Equivariant Learning for Compositional VQA

Zhou Du, Zhaoquan Yuan, Xiao Wu +1

Compositional visual question answering (VQA) represents a challenging yet fundamental task that requires models to comprehend novel combinations of previously learned concepts. Th…

cs.CV2026★ 1 cited

Adaptive 3D Convolution for Remote Sensing Image Fusion

Siran Peng, Xiangyu Zhu, Shang-Qi Deng +2

Remote sensing image fusion aims to create a high-resolution multi/hyper-spectral image from a high-resolution image with limited spectral information and a low-resolution image wi…