From the 1 of 12 linked papers with an AI index.
8 papers · 1 filter
Semantically Calibrated Evidence Composition for CT Vision-Language Learning
Guoliang You, Haifan Gong, Xiaomeng Chu
Learning transferable representations from CT-report pairs requires combining whole-volume context with anatomy-specific evidence. Existing methods typically emphasize either globa…
Learning How Much, Not Just What: Cross-Patient Burden Order for CT Vision-Language Pretraining
Guoliang You, Haifan Gong, Xiaomeng Chu
Volumetric CT vision-language pretraining learns 3D representations from scan-report pairs, but global and anatomy-aware objectives supervise only correspondence: they establish wh…
Fine-Grained Vision-Language Pretraining with Organ-Conditioned Pattern Tokens for CT Understanding
Guoliang You, Xiaomeng Chu
The paper introduces OCP-CT, a framework that aligns organ‑conditioned radiological pattern tokens between CT scans and radiology reports using a mixture‑of‑experts and contrastive…
RaCFormer: Towards High-Quality 3D Object Detection via Query-based Radar-Camera Fusion
Xiaomeng Chu, Jiajun Deng, Guoliang You +3
We propose Radar-Camera fusion transformer (RaCFormer) to boost the accuracy of 3D object detection by the following insight. The Radar-Camera fusion in outdoor 3D scene perception…
S3R-GS: Streamlining the Pipeline for Large-Scale Street Scene Reconstruction
Guangting Zheng, Jiajun Deng, Xiaomeng Chu +3
Recently, 3D Gaussian Splatting (3DGS) has reshaped the field of photorealistic 3D reconstruction, achieving impressive rendering quality and speed. However, when applied to large-…
Perception Helps Planning: Facilitating Multi-Stage Lane-Level Integration via Double-Edge Structures
Guoliang You, Xiaomeng Chu, Yifan Duan +6
When planning for autonomous driving, it is crucial to consider essential traffic elements such as lanes, intersections, traffic regulations, and dynamic agents. However, they are…