7 papers
Latent-Frequency Validity: Fast Spectral Editing with Screened Video-VAE Transfer Operators
Bowen Xue, Jiafeng Xiong, Xin Quan
Direct spectral editing in video-VAE latents can control noise, flicker, smoothness, and frequency content without a decode--filter--reencode pass. However, video VAEs may redistri…
FourTune: Towards Fully 4-Bit Efficient Post-Training for Diffusion Models
Bowen Xue, Zihan Min, Xingyang Li +8
Diffusion models have become a dominant paradigm for high-quality generative modeling, while post-training is essential for adapting them to diverse downstream applications. Howeve…
VideoNeuMat: Neural Material Extraction from Generative Video Models
Bowen Xue, Saeed Hadadan, Zheng Zeng +3
Creating photorealistic materials for 3D rendering requires exceptional artistic skill. Generative models for materials could help, but are currently limited by the lack of high-qu…
Stand-In: A Lightweight and Plug-and-Play Identity Control for Video Generation
Bowen Xue, Zheng-Peng Duan, Qixin Yan +6
Generating high-fidelity human videos that match user-specified identities is important yet challenging in the field of generative AI. Existing methods often rely on an excessive n…
PBR-Inspired Controllable Diffusion for Image Generation
Bowen Xue, Giuseppe Claudio Guarnera, Shuang Zhao +1
Despite recent advances in text-to-image generation, controlling geometric layout and PBR material properties in synthesized scenes remains challenging. We present a pipeline that…
OneSearch: A Preliminary Exploration of the Unified End-to-End Generative Framework for E-commerce Search
Ben Chen, Xian Guo, Siyuan Wang +25
Traditional e-commerce search systems employ multi-stage cascading architectures (MCA) that progressively filter items through recall, pre-ranking, and ranking stages. While effect…