4 papers · 1 filter
ExPLoRA: Parameter-Efficient Extended Pre-Training to Adapt Vision Transformers under Domain Shifts
Samar Khanna, Medhanie Irgau, David B. Lobell +1
Parameter-efficient fine-tuning (PEFT) techniques such as low-rank adaptation (LoRA) can effectively adapt large pre-trained foundation models to downstream tasks using only a smal…
TEOChat: A Large Vision-Language Assistant for Temporal Earth Observation Data
Jeremy Andrew Irvin, Emily Ruoyu Liu, Joyce Chuyi Chen +5
Large vision and language assistants have enabled new capabilities for interpreting natural images. These approaches have recently been adapted to earth observation data, but they…
DiffusionSat: A Generative Foundation Model for Satellite Imagery
Samar Khanna, Patrick Liu, Linqi Zhou +5
Diffusion models have achieved state-of-the-art results on many modalities including images, speech, and video. However, existing models are not tailored to support remote sensing…
SpotNet: An Image Centric, Lidar Anchored Approach To Long Range Perception
Louis Foucard, Samar Khanna, Yi Shi +4
In this paper, we propose SpotNet: a fast, single stage, image-centric but LiDAR anchored approach for long range 3D object detection. We demonstrate that our approach to LiDAR/ima…