2 papers
cs.CV2026
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT
Alaa Asfour, Christopher Indris, Leihan Chen +2
Large-scale 3D vision-language models (VLMs) like LLaVA-3D offer strong spatial reasoning but are difficult to deploy due to high computational costs. We propose a knowledge distil…
cs.CV2025
Tracking Moose using Aerial Object Detection
Christopher Indris, Raiyan Rahman, Goetz Bramesfeld +1
Aerial wildlife tracking is critical for conservation efforts and relies on detecting small objects on the ground below the aircraft. It presents technical challenges: crewed aircr…