collaborators

5 papers

cs.CV2026

Lift, Associate, and Fuse: A Decision-Centric Framework for 2D-to-3D Foundation Model Transfer

Wentao Sun, Yiping Chen, John S. Zelek +1

Methods that transfer predictions from two-dimensional foundation models into three-dimensional segmentation are commonly grouped by task or representation. Those groupings obscure…

cs.CV2026

CDSeg: A Renderable Gaussian Carrier for Image-to-3D Label Transfer

Wentao Sun, Yiping Chen, Zhengsen Xu +2

Modern image models provide strong cues about \emph{what} should be segmented in each view, but their masks do not by themselves determine \emph{where} those labels should persist…

cs.CV2026

Out-of-Length Scene Text Recognition: A Two-Axis Diagnosis and a Training-Free Fix

Zobeir Raisi, John Zelek

Scene Text Recognition (STR) models are trained almost exclusively on word crops of at most 25 characters, yet real deployments (signage, product labels, dense captions) require re…

cs.CV2026

SAGOnline: Segment Any Gaussians Online

Wentao Sun, Quanyun Wu, Hanqing Xu +7

3D Gaussian Splatting has emerged as a powerful paradigm for explicit 3D scene representation, yet achieving efficient and consistent 3D segmentation remains challenging. Existing…

cs.CV2025

PointGauss: Point Cloud-Guided Multi-Object Segmentation for Gaussian Splatting

Wentao Sun, Hanqing Xu, Quanyun Wu +5

We introduce PointGauss, a novel point cloud-guided framework for real-time multi-object segmentation in Gaussian Splatting representations. Unlike existing methods that suffer fro…