most citedCVP: Central-Peripheral Vision-Inspired Multimodal Model for Spatial Reasoning

1 citations · 1 across the 4 of their papers we have counts for

collaborators

8 papers

cs.CV2026

Soft Tail-dropping for Adaptive Visual Tokenization

Zeyuan Chen, Kai Zhang, Zhuowen Tu +1

We present Soft Tail-dropping Adaptive Tokenizer (STAT), a 1D discrete visual tokenizer that adaptively chooses the number of output tokens per image according to its structural co…

cs.CV20251 cited

CVP: Central-Peripheral Vision-Inspired Multimodal Model for Spatial Reasoning

Zeyuan Chen, Xiang Zhang, Haiyang Xu +2

We present a central-peripheral vision-inspired framework (CVP), a simple yet effective multimodal model for spatial reasoning that draws inspiration from the two types of human vi…

cs.CV2025

Gaussian Swaying: Surface-Based Framework for Aerodynamic Simulation with 3D Gaussians

Hongru Yan, Xiang Zhang, Zeyuan Chen +2

Branches swaying in the breeze, flags rippling in the wind, and boats rocking on the water all show how aerodynamics shape natural motion -- an effect crucial for realism in vision…

cs.GR2025

C3Editor: Achieving Controllable Consistency in 2D Model for 3D Editing

Zeng Tao, Zheng Ding, Zeyuan Chen +3

Existing 2D-lifting-based 3D editing methods often encounter challenges related to inconsistency, stemming from the lack of view-consistent 2D editing models and the difficulty of…

cs.CV2025

OverLayBench: A Benchmark for Layout-to-Image Generation with Dense Overlaps

Bingnan Li, Chen-Yu Wang, Haiyang Xu +7

Despite steady progress in layout-to-image generation, current methods still struggle with layouts containing significant overlap between bounding boxes. We identify two primary ch…

cs.CV2025

Exploring the Equivalence of Closed-Set Generative and Real Data Augmentation in Image Classification

Haowen Wang, Guowei Zhang, Xiang Zhang +4

In this paper, we address a key scientific problem in machine learning: Given a training set for an image classification task, can we train a generative model on this dataset to en…