3 papers
cs.RO2026
Immiscible Diffusion Policy: Preserving Multimodal Robot Actions through Label-Free Noise Assignment
Xiao Zhang, Yuxin Chen, Zhixuan Liang +5
When diffusion policies were first introduced, they were expected to recover multi-modal action distributions. However, we find this expectation does not always hold, as diffusion…
cs.CV2026
Adaptive Visual Token Reduction for Accelerated Image Understanding
Seyoung Jeong, Jong Pil Yun, Sang Jun Lee
Large Vision-Language Models achieve strong VQA performance, but processing high-resolution, information-rich images requires substantial computation, motivating visual token reduc…
cs.CV2026
M2P-AD: Memory-to-Prototype Learning with Boundary-aware Score Refinement for 3D Anomaly Detection
Seyoung Jeong, Jong Pil Yun, Sang Jun Lee
3D anomaly detection has recently emerged as an important research topic in computer vision. Although existing methods have achieved high performance, excessive anomaly responses i…