10 citations · 14 across the 11 of their papers we have counts for
9 papers · 1 filter
MCPNet: An Interpretable Classifier via Multi-Level Concept Prototypes
Bor-Shiun Wang, Chien-Yi Wang, Wei-Chen Chiu
Recent advancements in post-hoc and inherently interpretable methods have markedly enhanced the explanations of black box classifier models. These methods operate either through po…
MENTOR: Multilingual tExt detectioN TOward leaRning by analogy
Hsin-Ju Lin, Tsu-Chun Chung, Ching-Chun Hsiao +3
Text detection is frequently used in vision-based mobile robots when they need to interpret texts in their surroundings to perform a given task. For instance, delivery robots in mu…
Improving Robustness for Joint Optimization of Camera Poses and Decomposed Low-Rank Tensorial Radiance Fields
Bo-Yu Cheng, Wei-Chen Chiu, Yu-Lun Liu
In this paper, we propose an algorithm that allows joint refinement of camera pose and scene geometry represented by decomposed low-rank tensor, using only 2D images as supervision…
Skin the sheep not only once: Reusing Various Depth Datasets to Drive the Learning of Optical Flow
Sheng-Chi Huang, Wei-Chen Chiu
Optical flow estimation is crucial for various applications in vision and robotics. As the difficulty of collecting ground truth optical flow in real-world scenarios, most of the e…
Transformer-based Image Compression with Variable Image Quality Objectives
Chia-Hao Kao, Yi-Hsin Chen, Cheng Chien +2
This paper presents a Transformer-based image compression system that allows for a variable image quality objective according to the user's preference. Optimizing a learned codec f…
Multimodal Prompting with Missing Modalities for Visual Recognition
Yi-Lun Lee, Yi-Hsuan Tsai, Wei-Chen Chiu +1
In this paper, we tackle two challenges in multimodal learning for visual recognition: 1) when missing-modality occurs either during training or testing in real-world situations; a…