126 citations · 191 across the 37 of their papers we have counts for
5 papers · 1 filter
SE-MoLoRA: Shared-Expert LoRA Adapters for Domain-Specific Photographic Assessment
Bishwash Khanal, Anlan Zhang, Sasu Tarkoma +2
Vision-language models can describe images fluently, but they often fail to provide actionable photographic critique because semantic content and aesthetic judgment remain entangle…
Exponentially Weighted Instance-Aware Repeat Factor Sampling for Long-Tailed Object Detection Model Training in Unmanned Aerial Vehicles Surveillance Scenarios
Taufiq Ahmed, Abhishek Kumar, Constantino Álvarez Casado +5
Object detection models often struggle with class imbalance, where rare categories appear significantly less frequently than common ones. Existing sampling-based rebalancing strate…
From Pixels to Progress: Generating Road Network from Satellite Imagery for Socioeconomic Insights in Impoverished Areas
Yanxin Xi, Yu Liu, Zhicheng Liu +3
The Sustainable Development Goals (SDGs) aim to resolve societal challenges, such as eradicating poverty and improving the lives of vulnerable populations in impoverished areas. Th…
A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming
Pengyuan Zhou, Lin Wang, Zhi Liu +4
This paper offers an insightful examination of how currently top-trending AI technologies, i.e., generative artificial intelligence (Generative AI) and large language models (LLMs)…
A Satellite Imagery Dataset for Long-Term Sustainable Development in United States Cities
Yanxin Xi, Yu Liu, Tong Li +5
Cities play an important role in achieving sustainable development goals (SDGs) to promote economic growth and meet social needs. Especially satellite imagery is a potential data s…