2 papers
cs.CV2026
Spatial Temporal Synergy: Balancing Change and Invariance in Text Driven 3D Human Motion Editing
Shaohui Lin, Zhenwu Shi, Jingyu Gong +5
Text-driven human motion editing aims to modify existing motion sequences according to natural language instructions while maintaining the structural consistency of the original mo…
cs.CV2026
CLIP-Map: Structured Matrix Mapping for Parameter-Efficient CLIP Compression
Kangjie Zhang, Wenxuan Huang, Xin Zhou +9
Contrastive Language-Image Pre-training (CLIP) has achieved widely applications in various computer vision tasks, e.g., text-to-image generation, Image-Text retrieval and Image cap…