Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
GTA-Net: Cooperative Game Theory for Vision-Language Alignment in Chest X-Ray Report Generation
Saif ur Rehman Khan, Imad Ahmed Waqar, Sebastian Vollmer +2
Automated chest X-ray report generation requires precise cross-modal grounding to ensure clinically reliable descriptions. However, existing vision-language models rely on implicit…
cs.CV2024
DistillGrasp: Integrating Features Correlation with Knowledge Distillation for Depth Completion of Transparent Objects
Yiheng Huang, Junhong Chen, Nick Michiels +3
Due to the visual properties of reflection and refraction, RGB-D cameras cannot accurately capture the depth of transparent objects, leading to incomplete depth maps. To fill in th…