activity
20242026
collaborators

5 papers

cs.RO2026

Causality-Aware End-to-End Autonomous Driving via Ego-Centric Joint Scene Modeling

Seokha Moon, Minseung Lee, Joon Seo +2

End-to-end autonomous driving, which bypasses traditional modular pipelines by directly predicting future trajectories from sensor inputs, has recently achieved substantial progres…

cs.CV2026

Focus, Don't Prune: Identifying Instruction-Relevant Regions for Information-Rich Image Understanding

Mincheol Kwon, Minseung Lee, Seonga Choi +7

Large Vision-Language Models (LVLMs) have shown strong performance across various multimodal tasks by leveraging the reasoning capabilities of Large Language Models (LLMs). However…

cs.CV2025

Image-Guided Semantic Pseudo-LiDAR Point Generation for 3D Object Detection

Minseung Lee, Seokha Moon, Seung Joon Lee +2

In autonomous driving scenarios, accurate perception is becoming an even more critical task for safe navigation. While LiDAR provides precise spatial data, its inherent sparsity ma…

cs.CV2025

Local Representative Token Guided Merging for Text-to-Image Generation

Min-Jeong Lee, Hee-Dong Kim, Seong-Whan Lee

Stable diffusion is an outstanding image generation model for text-to-image, but its time-consuming generation process remains a challenge due to the quadratic complexity of attent…

cs.AI2024

Slot State Space Models

Jindong Jiang, Fei Deng, Gautam Singh +2

Recent State Space Models (SSMs) such as S4, S5, and Mamba have shown remarkable computational benefits in long-range temporal dependency modeling. However, in many sequence modeli…