Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Prune Once: Retraining-Free Task-Agnostic Pruning for Vision-Language Models
Minseok Kang, Hyunwoo Kim, Chanyoung Kim +3
Vision-language models (VLMs) have achieved remarkable generalization across diverse multimodal tasks through large-scale pre-training, yet their rapidly increasing computational a…
cs.CV2025
GUIDE-CoT: Goal-driven and User-Informed Dynamic Estimation for Pedestrian Trajectory using Chain-of-Thought
Sungsik Kim, Janghyun Baek, Jinkyu Kim +1
While Large Language Models (LLMs) have recently shown impressive results in reasoning tasks, their application to pedestrian trajectory prediction remains challenging due to two k…
cs.CV2024
Learning Temporal Cues by Predicting Objects Move for Multi-camera 3D Object Detection
Seokha Moon, Hongbeen Park, Jungphil Kwon +2
In autonomous driving and robotics, there is a growing interest in utilizing short-term historical data to enhance multi-camera 3D object detection, leveraging the continuous and c…