2 papers
cs.CV2026
Pareto-Guided Optimal Transport for Multi-Reward Alignment
Ying Ba, Tianyu Zhang, Mohan Zhou +5
Text-to-image generation models have achieved remarkable progress in preference optimization, yet achieving robust alignment across diverse reward models remains a significant chal…
cs.CV2025
V2Flow: Unifying Visual Tokenization and Large Language Model Vocabularies for Autoregressive Image Generation
Guiwei Zhang, Tianyu Zhang, Mohan Zhou +2
We propose V2Flow, a novel tokenizer that produces discrete visual tokens capable of high-fidelity reconstruction, while ensuring structural and latent distribution alignment with…