2 papers
cs.CV2024
OneChart: Purify the Chart Structural Extraction via One Auxiliary Token
Jinyue Chen, Lingyu Kong, Haoran Wei +6
Chart parsing poses a significant challenge due to the diversity of styles, values, texts, and so forth. Even advanced large vision-language models (LVLMs) with billions of paramet…
cs.CV2023
Contrastive Grouping with Transformer for Referring Image Segmentation
Jiajin Tang, Ge Zheng, Cheng Shi +1
Referring image segmentation aims to segment the target referent in an image conditioning on a natural language expression. Existing one-stage methods employ per-pixel classificati…