8 papers
DiRL: An Efficient Post-Training Framework for Diffusion Language Models
Ying Zhu, Jiaxin Wan, Xiaoran Liu +7
Diffusion Language Models (dLLMs) have emerged as promising alternatives to Auto-Regressive (AR) models. While recent efforts have validated their pre-training potential and accele…
MPJudge: Towards Perceptual Assessment of Music-Induced Paintings
Shiqi Jiang, Tianyi Liang, Huayuan Ye +2
Music induced painting is a unique artistic practice, where visual artworks are created under the influence of music. Evaluating whether a painting faithfully reflects the music th…
IFDECORATOR: Wrapping Instruction Following Reinforcement Learning with Verifiable Rewards
Xu Guo, Tianyi Liang, Tong Jian +6
Reinforcement Learning with Verifiable Rewards (RLVR) improves instruction following capabilities of large language models (LLMs), but suffers from training inefficiency due to ina…
Sel3DCraft: Interactive Visual Prompts for User-Friendly Text-to-3D Generation
Nan Xiang, Tianyi Liang, Haiwen Huang +6
Text-to-3D (T23D) generation has transformed digital content creation, yet remains bottlenecked by blind trial-and-error prompting processes that yield unpredictable results. While…
Music2Palette: Emotion-aligned Color Palette Generation via Cross-Modal Representation Learning
Jiayun Hu, Yueyi He, Tianyi Liang +2
Emotion alignment between music and palettes is crucial for effective multimedia content, yet misalignment creates confusion that weakens the intended message. However, existing me…
RouteWinFormer: A Route-Window Transformer for Middle-range Attention in Image Restoration
Qifan Li, Tianyi Liang, Xingtao Wang +1
Transformer models have recently garnered significant attention in image restoration due to their ability to capture long-range pixel dependencies. However, long-range attention of…