3 papers
cs.CV2026
Attribute Token Arithmetic: Disentangled and Continuous Semantic Control for Visual Autoregressive Models
Xindi Yang, Yicheng Wu, Cheng Zhang +2
Autoregressive text-to-image generation has recently achieved remarkable progress, offering high-fidelity synthesis via a unified generative framework. However, fine-grained semant…
cs.CV2025
OpenView: Empowering MLLMs with Out-of-view VQA
Qixiang Chen, Cheng Zhang, Chi-Wing Fu +2
Recent multimodal large language models (MLLMs) show great potential in natural image understanding. Yet, they perform well, mainly on reasoning in-view contents within the image f…
cs.CV2025
ASAP-Textured Gaussians: Enhancing Textured Gaussians with Adaptive Sampling and Anisotropic Parameterization
Meng Wei, Cheng Zhang, Jianmin Zheng +2
Recent advances have equipped 3D Gaussian Splatting with texture parameterizations to capture spatially varying attributes, improving the performance of both appearance modeling an…