3 papers
cs.RO2026
ProgVLA: Progress-Aware Robot Manipulation Skill Learning
Seungsu Kim, Jinyoung Choi, Seungmin Baek +1
We present ProgVLA, a compact vision-language-action (VLA) model designed for reliable robot manipulation under tight compute and memory budgets. The model specifically focuses on…
cs.CV2025
Difficulty-Aware Label-Guided Denoising for Monocular 3D Object Detection
Soyul Lee, Seungmin Baek, Dongbo Min
Monocular 3D object detection is a cost-effective solution for applications like autonomous driving and robotics, but remains fundamentally ill-posed due to inherently ambiguous de…
cs.CV2025
TADFormer : Task-Adaptive Dynamic Transformer for Efficient Multi-Task Learning
Seungmin Baek, Soyul Lee, Hayeon Jo +2
Transfer learning paradigm has driven substantial advancements in various vision tasks. However, as state-of-the-art models continue to grow, classical full fine-tuning often becom…