Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Dual-Stream Diffusion for World-Model Augmented Vision-Language-Action Model
John Won, Kyungmin Lee, Huiwon Jang +2
Augmenting vision-language-action models (VLAs) with world models is promising for robotic policy learning but faces challenges in jointly predicting states and actions due to the…
cs.CV2025
CLIP Meets Diffusion: A Synergistic Approach to Anomaly Detection
Byeongchan Lee, John Won, Seunghyun Lee +1
Anomaly detection is a complex problem due to the ambiguity in defining anomalies, the diversity of anomaly types (e.g., local and global defect), and the scarcity of training data…