activity
20242026
collaborators

8 papers

cs.CV2026

FedSmoothLoRA: Toward Smoother and Faster Convergence in Federated Low-Rank Adaptation

Zehao Wang, Guanglei Yang, Yihan Zeng +4

Federated fine-tuning of foundation models with Low-Rank Adaptation (LoRA) provides an efficient solution for reducing communication and computation costs while preserving data loc…

cs.CV2026

Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation

Zihao Wang, Yuxiang Wei, Xinpeng Zhou +5

Text-to-image generation has advanced rapidly, yet it still struggles to capture the nuanced user preferences. Existing approaches typically rely on multimodal large language model…

cs.CV2025

PhysChoreo: Physics-Controllable Video Generation with Part-Aware Semantic Grounding

Haoze Zhang, Tianyu Huang, Zichen Wan +4

While recent video generation models have achieved significant visual fidelity, they often suffer from the lack of explicit physical controllability and plausibility. To address th…

cs.GR2025

Pseudo-Label Guided Real-World Image De-weathering: A Learning Framework with Imperfect Supervision

Heming Xu, Xiaohui Liu, Zhilu Zhang +3

Real-world image de-weathering aims at removingvarious undesirable weather-related artifacts, e.g., rain, snow,and fog. To this end, acquiring ideal training pairs is crucial.Exist…

cs.CV2024

UniRestorer: Universal Image Restoration via Adaptively Estimating Image Degradation at Proper Granularity

Jingbo Lin, Zhilu Zhang, Wenbo Li +4

Recently, considerable progress has been made in all-in-one image restoration. Generally, existing methods can be degradation-agnostic or degradation-aware. However, the former are…

cs.CV2024

VitaGlyph: Vitalizing Artistic Typography with Flexible Dual-branch Diffusion Models

Kailai Feng, Yabo Zhang, Haodong Yu +4

Artistic typography is a technique to visualize the meaning of input character in an imaginable and readable manner. With powerful text-to-image diffusion models, existing methods…