activity
20232026
most citedContext Does Matter: End-to-end Panoptic Narrative Grounding with Deformable Attention Refined Matching Network

1 citations · 1 across the 6 of their papers we have counts for

collaborators

5 papers

cs.CL2026

Reactivating Test-Time Scaling for Plane Geometry Problem Solving

Xiaoqiang Kang, Shengen Wu, Maizhen Ning +5

Plane geometry problem (PGP) solving has become a critical benchmark for multimodal reasoning because it requires accurate visual perception and precise multi-step symbolic deducti…

cs.CV2025

Hyperbolic Structured Classification for Robust Single Positive Multi-label Learning

Yiming Lin, Shang Wang, Junkai Zhou +3

Single Positive Multi-Label Learning (SPMLL) addresses the challenging scenario where each training sample is annotated with only one positive label despite potentially belonging t…

cs.CL2025

Can GRPO Boost Complex Multimodal Table Understanding?

Xiaoqiang Kang, Shengen Wu, Zimu Wang +7

Existing table understanding methods face challenges due to complex table structures and intricate logical reasoning. While supervised finetuning (SFT) dominates existing research,…

cs.CV2025

The Demon is in Ambiguity: Revisiting Situation Recognition with Single Positive Multi-Label Learning

Yiming Lin, Yuchen Niu, Shang Wang +3

Context recognition (SR) is a fundamental task in computer vision that aims to extract structured semantic summaries from images by identifying key events and their associated enti…

cs.CL2024

Template-Driven LLM-Paraphrased Framework for Tabular Math Word Problem Generation

Xiaoqiang Kang, Zimu Wang, Xiaobo Jin +3

Solving tabular math word problems (TMWPs) has become a critical role in evaluating the mathematical reasoning ability of large language models (LLMs), where large-scale TMWP sampl…