3 citations · 4 across the 4 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
FrogDogNet: Fourier frequency Retained visual prompt Output Guidance for Domain Generalization of CLIP in Remote Sensing
Hariseetharam Gunduboina, Muhammad Haris Khan, Biplab Banerjee
In recent years, large-scale vision-language models (VLMs) like CLIP have gained attention for their zero-shot inference using instructional text prompts. While these models excel…
cs.CV2025★ 1 cited
ASTRA: A Scene-aware TRAnsformer-based model for trajectory prediction
Izzeddin Teeti, Aniket Thomas, Munish Monga +5
We present ASTRA (A} Scene-aware TRAnsformer-based model for trajectory prediction), a light-weight pedestrian trajectory forecasting model that integrates the scene context, spati…