1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2025
Benchmarking Multimodal Mathematical Reasoning with Explicit Visual Dependency
Zhikai Wang, Jiashuo Sun, Wenqi Zhang +4
Recent advancements in Large Vision-Language Models (LVLMs) have significantly enhanced their ability to integrate visual and linguistic information, achieving near-human proficien…
math.CO2024
Hamiltonian cycles passing through matchings in -ary -cubes
Baolai Liao, Fan Wang
As we all know, the -ary -cube is a highly efficient interconnect network topology structure. It is also a concept of great significance, with a broad range of applications s…
cs.CV2023★ 1 cited
Dynamic Token-Pass Transformers for Semantic Segmentation
Yuang Liu, Qiang Zhou, Jing Wang +3
Vision transformers (ViT) usually extract features via forwarding all the tokens in the self-attention layers from top to toe. In this paper, we introduce dynamic token-pass vision…