1 paper
Haoran Wang, Chaofan Ma, Ran Yi +1
Despite recent advances in unified multimodal models for multi-reference image generation, existing benchmarks remain organized around predefined task types (e.g., "subject composi…