From the 1 of 9 linked papers with an AI index.
9 papers
R-SLPR: Region-based Small-to-Large Point-cloud Registration with Contrastive Learning
Yusen Wan, Zeyuan Chen, Qianshi Zou +1
The paper introduces R‑SLPR, a three‑stage framework that registers a small, partial point cloud to a much larger reference by proposing regions, matching them with contrastive lea…
Soft Tail-dropping for Adaptive Visual Tokenization
Zeyuan Chen, Kai Zhang, Zhuowen Tu +1
We present Soft Tail-dropping Adaptive Tokenizer (STAT), a 1D discrete visual tokenizer that adaptively chooses the number of output tokens per image according to its structural co…
CVP: Central-Peripheral Vision-Inspired Multimodal Model for Spatial Reasoning
Zeyuan Chen, Xiang Zhang, Haiyang Xu +2
We present a central-peripheral vision-inspired framework (CVP), a simple yet effective multimodal model for spatial reasoning that draws inspiration from the two types of human vi…
Gaussian Swaying: Surface-Based Framework for Aerodynamic Simulation with 3D Gaussians
Hongru Yan, Xiang Zhang, Zeyuan Chen +2
Branches swaying in the breeze, flags rippling in the wind, and boats rocking on the water all show how aerodynamics shape natural motion -- an effect crucial for realism in vision…
C3Editor: Achieving Controllable Consistency in 2D Model for 3D Editing
Zeng Tao, Zheng Ding, Zeyuan Chen +3
Existing 2D-lifting-based 3D editing methods often encounter challenges related to inconsistency, stemming from the lack of view-consistent 2D editing models and the difficulty of…
OverLayBench: A Benchmark for Layout-to-Image Generation with Dense Overlaps
Bingnan Li, Chen-Yu Wang, Haiyang Xu +7
Despite steady progress in layout-to-image generation, current methods still struggle with layouts containing significant overlap between bounding boxes. We identify two primary ch…