3 papers
cs.CV2025
Graph Integrated Multimodal Concept Bottleneck Model
Jiakai Lin, Jinchang Zhang, Guoyu Lu
With growing demand for interpretability in deep learning, especially in high stakes domains, Concept Bottleneck Models (CBMs) address this by inserting human understandable concep…
cs.CV2025
Vision-Language Embodiment for Monocular Depth Estimation
Jinchang Zhang, Guoyu Lu
Depth estimation is a core problem in robotic perception and vision tasks, but 3D reconstruction from a single image presents inherent uncertainties. Current depth estimation model…
cs.CV2025
Keypoint Detection and Description for Raw Bayer Images
Jiakai Lin, Jinchang Zhang, Guoyu Lu
Keypoint detection and local feature description are fundamental tasks in robotic perception, critical for applications such as SLAM, robot localization, feature matching, pose est…