activity
20242026
collaborators
Showing 2026 · cs.CVShow all

14 papers · 2 filters

cs.CV2026

Pixel-Space Diffusion via Observation Operators

Shaojie Guo, Lichen Ma, Haoyang Tong +8

Pixel-space diffusion models directly model image distributions but remain difficult to optimize. Recent methods alleviate this challenge through target reparameterization, while s…

cs.CV2026

PosterText: Towards Unified Visual Text Generation and Editing for E-commerce Poster

Xiaoan Liu, Lichen Ma, Zipeng Guo +12

Automated e-commerce poster design requires both high-quality poster generation and flexible editing of existing designs. However, most existing methods either target end-to-end po…

cs.CV2026

TransAnyText: Translating Arbitrary Text in E-commerce Images via Structured Visual Generation

Xiaoan Liu, Lichen Ma, Zipeng Guo +13

Cross-border e-commerce image translation is essential for global retail, where product images, banners, and detail pages need to be produced in different languages. Existing metho…

cs.CV2026

Energy-Guided Flow Matching

Haoyang Tong, Yu He, Fang Li +6

Pixel-space generative models bypass lossy latent compression, yet necessitate joint learning of global structure and fine-grained details in a high-dimensional space. Standard flo…

cs.CV2026

iFAN: Inference-Aware Learning for Plain Mask Transformers

Fang Li, Yu He, Haoyang Tong +7

Query-based mask transformers assemble segmentation outputs through pixel-wise competition among query predictions of the final layer, yet this inference process is not explicitly…

cs.CV2026

Match One, Learn with Graph: One-to-Graph Query Collaboration with Backward Sharing for Object Detection

Wenxiao Fan, Jingling Fu, Luohang Liu +8

One-to-one (O2O) matching enables Detection Transformers (DETRs) to perform end-to-end set prediction by assigning each object to a single positive query. However, the strongest cl…