3 papers
cs.CV2026
OctoSense: Self-Supervised Learning for Multimodal Robot Perception
Anthony Bisulco, Jeremy Wang, Kostas Daniilidis +2
We present OctoSense, an open-source sensor platform with stereo RGB and event cameras, LiDAR, a thermal camera, an inertial measurement unit, RTK-corrected global positioning syst…
cs.CV2025
From Linearity to Non-Linearity: How Masked Autoencoders Capture Spatial Correlations
Anthony Bisulco, Rahul Ramesh, Randall Balestriero +1
Masked Autoencoders (MAEs) have emerged as a powerful pretraining technique for vision foundation models. Despite their effectiveness, they require extensive hyperparameter tuning…
cs.CV2025
Many Perception Tasks are Highly Redundant Functions of their Input Data
Rahul Ramesh, Anthony Bisulco, Ronald W. DiTullio +4
We show that many perception tasks, from visual recognition, semantic segmentation, optical flow, depth estimation to vocalization discrimination, are highly redundant functions of…