From the 1 of 7 linked papers with an AI index.
7 papers
MORFEO control strategy
Guido Agapito, Lorenzo Busoni, Cédric Plantet +7
The ESO Extremely Large Telescope (ELT) will offer unprecedented sensitivity and resolution in the near-infrared, marking a new era for ground-based astronomy. Among its key imagin…
Slot2Text: Object-Centric Visual Tokenization for Efficient and Spatially Traceable Surgical MLLMs
Guiqiu Liao, Matjaz Jogan, Daniel A. Hashimoto
Multimodal large language models (MLLM) for surgical scene understanding typically inject hundreds of dense visual tokens into a language model, leading to costly inference and lim…
Active Learning for Efficient Annotation of Surgical Videos with Weak Supervision
Manasa Dendukuri, Matjaz Jogan, Daniel A. Hashimoto +1
The paper presents a human‑in‑the‑loop framework that combines active learning with dual‑loss weak supervision to cut the effort needed for annotating laparoscopic video frames, en…
DenseTRF: Texture-Aware Unsupervised Representation Adaptation for Surgical Scene Dense Prediction
Guiqiu Liao, Matjaž Jogan, Daniel A. Hashimoto
Dense prediction tasks in surgical computer vision, such as segmentation and surgical zone prediction, can provide valuable guidance for laparoscopic and robotic surgery. However,…
Slot-BERT: Self-supervised Object Discovery in Surgical Video
Guiqiu Liao, Matjaz Jogan, Marcel Hussing +5
Object-centric slot attention is a powerful framework for unsupervised learning of structured and explainable representations that can support reasoning about objects and actions,…
FORLA: Federated Object-centric Representation Learning with Slot Attention
Guiqiu Liao, Matjaz Jogan, Eric Eaton +1
Learning efficient visual representations across heterogeneous unlabeled datasets remains a central challenge in federated learning. Effective federated representations require fea…