From the 1 of 11 linked papers with an AI index.
11 papers
CoVStream: Edge-Cloud Collaboration for Understanding of Long Video Streams
Xu Liu, Guikun Chen, Zihao Yan +2
The paper introduces CoVStream, an edge‑cloud system that compresses raw video into compact visual features and captions on the device, sends them to the cloud for graph‑based reas…
Moving Beyond Diffusion: Hierarchy-to-Hierarchy Autoregression for fMRI-to-Image Reconstruction
Xu Zhang, Ruijie Quan, Wenguan Wang +1
Reconstructing visual stimuli from fMRI signals is a central challenge bridging machine learning and neuroscience. Recent diffusion-based methods typically map fMRI activity to a s…
Uncertainty-Aware Gaussian Map for Vision-Language Navigation
Jianzhe Gao, Rui Liu, Yuxuan Xu +6
Vision-Language Navigation (VLN) requires an agent to navigate 3D environments following natural language instructions. During navigation, existing agents commonly encounter percep…
Clinically-Grounded Counterfactual Reasoning for Medical Video Diagnosis
Jianzhe Gao, Churan Wang, Weiyi Zhang +5
Medical video diagnosis involves inferring clinical decisions from dynamic tissue responses throughout examination processes. Existing methods rely on an end-to-end learning paradi…
A Survey on 3D Gaussian Splatting
Guikun Chen, Wenguan Wang
3D Gaussian splatting (GS) has emerged as a transformative technique in radiance fields. Unlike mainstream implicit neural models, 3D GS uses millions of learnable 3D Gaussians for…
A Survey of World Models for Autonomous Driving
Tuo Feng, Wenguan Wang, Yi Yang
Recent breakthroughs in autonomous driving have been propelled by advances in robust world modeling, fundamentally transforming how vehicles interpret dynamic scenes and execute sa…