Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
U2-BENCH: Benchmarking Large Vision-Language Models on Ultrasound Understanding
Anjie Le, Henan Liu, Yue Wang +18
Ultrasound is a widely-used imaging modality critical to global healthcare, yet its interpretation remains challenging due to its varying image quality on operators, noises, and an…
cs.CV2024
Tri-modal Confluence with Temporal Dynamics for Scene Graph Generation in Operating Rooms
Diandian Guo, Manxi Lin, Jialun Pei +3
A comprehensive understanding of surgical scenes allows for monitoring of the surgical process, reducing the occurrence of accidents and enhancing efficiency for medical profession…
cs.CV2024
S^2Former-OR: Single-Stage Bi-Modal Transformer for Scene Graph Generation in OR
Jialun Pei, Diandian Guo, Jingyang Zhang +3
Scene graph generation (SGG) of surgical procedures is crucial in enhancing holistically cognitive intelligence in the operating room (OR). However, previous works have primarily r…