3 papers
cs.CV2026
How do Self-Supervised Remote Sensing Vision Models Transfer to Downstream Tasks?
Julia Romero, Qin Lv, Morteza Karimzadeh
Self-supervised geospatial foundation models (GeoFMs) learn transferable representations from remote sensing data, but their downstream behavior is difficult to characterize. We st…
cs.CV2025
Keystep Recognition using Graph Neural Networks
Julia Lee Romero, Kyle Min, Subarna Tripathi +1
We pose keystep recognition as a node classification task, and propose a flexible graph-learning framework for fine-grained keystep recognition that is able to effectively leverage…
cs.CV2025
Graph-Based Multimodal and Multi-view Alignment for Keystep Recognition
Julia Lee Romero, Kyle Min, Subarna Tripathi +1
Egocentric videos capture scenes from a wearer's viewpoint, resulting in dynamic backgrounds, frequent motion, and occlusions, posing challenges to accurate keystep recognition. We…