4 papers
Enabling Training-Free Text-Based Remote Sensing Segmentation
Jose Sosa, Danila Rukhovich, Anis Kacem +1
Recent advances in Vision Language Models (VLMs) and Vision Foundation Models (VFMs) have opened new opportunities for zero-shot text-guided segmentation of remote sensing imagery.…
Annotation Free Spacecraft Detection and Segmentation using Vision Language Models
Samet Hicsonmez, Jose Sosa, Dan Pineau +4
Vision Language Models (VLMs) have demonstrated remarkable performance in open-world zero-shot visual recognition. However, their potential in space-related applications remains la…
Motion Aware ViT-based Framework for Monocular 6-DoF Spacecraft Pose Estimation
Jose Sosa, Dan Pineau, Arunkumar Rathinam +2
Monocular 6-DoF pose estimation plays an important role in multiple spacecraft missions. Most existing pose estimation approaches rely on single images with static keypoint localis…
MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks
Jose Sosa, Danila Rukhovich, Anis Kacem +1
Multi-modal data in Earth Observation (EO) presents a huge opportunity for improving transfer learning capabilities when pre-training deep learning models. Unlike prior work that o…