1 citations · 2 across the 4 of their papers we have counts for
4 papers
Ocean-OCR: Towards General OCR Application via a Vision-Language Model
Song Chen, Xinyu Guo, Yadong Li +10
Multimodal large language models (MLLMs) have shown impressive capabilities across various domains, excelling in processing and understanding information from multiple modalities.…
Mono2Stereo: Monocular Knowledge Transfer for Enhanced Stereo Matching
Yuran Wang, Yingping Liang, Hesong Li +1
The generalization and performance of stereo matching networks are limited due to the domain gap of the existing synthetic datasets and the sparseness of GT labels in the real data…
Terrain Point Cloud Inpainting via Signal Decomposition
Yizhou Xie, Xiangning Xie, Yuran Wang +2
The rapid development of 3D acquisition technology has made it possible to obtain point clouds of real-world terrains. However, due to limitations in sensor acquisition technology…
Contributing Dimension Structure of Deep Feature for Coreset Selection
Zhijing Wan, Zhixiang Wang, Yuran Wang +3
Coreset selection seeks to choose a subset of crucial training samples for efficient learning. It has gained traction in deep learning, particularly with the surge in training data…