activity
20242026
collaborators

10 papers

cs.CV2026

A Creative Agent is Worth a 64-Token Template

Ruixiao Shi, Fu Feng, Yucheng Xie +3

Text-to-image (T2I) models have substantially improved image fidelity and prompt adherence, yet their creativity remains constrained by reliance on discrete natural language prompt…

cs.CV2026

Self-Supervised Weight Templates for Scalable Vision Model Initialization

Yucheng Xie, Fu Feng, Ruixiao Shi +3

The increasing scale and complexity of modern model parameters underscore the importance of pre-trained models. However, deployment often demands architectures of varying sizes, ex…

cs.LG2025

Knowledge Diversion for Efficient Morphology Control and Policy Transfer

Fu Feng, Ruixiao Shi, Yucheng Xie +3

Universal morphology control aims to learn a universal policy that generalizes across heterogeneous agent morphologies, with Transformer-based controllers emerging as a popular cho…

cs.CV2025

DivControl: Knowledge Diversion for Controllable Image Generation

Yucheng Xie, Fu Feng, Ruixiao Shi +3

Diffusion models have advanced from text-to-image (T2I) to image-to-image (I2I) generation by incorporating structured inputs such as depth maps, enabling fine-grained spatial cont…

cs.CV2025

FAD: Frequency Adaptation and Diversion for Cross-domain Few-shot Learning

Ruixiao Shi, Fu Feng, Yucheng Xie +2

Cross-domain few-shot learning (CD-FSL) requires models to generalize from limited labeled samples under significant distribution shifts. While recent methods enhance adaptability…

cs.CV2025

Distribution-Conditional Generation: From Class Distribution to Creative Generation

Fu Feng, Yucheng Xie, Xu Yang +2

Text-to-image (T2I) diffusion models are effective at producing semantically aligned images, but their reliance on training data distributions limits their ability to synthesize tr…