Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
KAN Text to Vision? The Exploration of Kolmogorov-Arnold Networks for Multi-Scale Sequence-Based Pose Animation from Sign Language Notation
Guanyi Du, Lintao Wang, Kun Hu +1
Sign language production from symbolic notation offers a scalable route to accessible sign animation. We present KANMultiSign, a multi-scale sequence generator that translates HamN…
cs.CV2024
SITransformer: Shared Information-Guided Transformer for Extreme Multimodal Summarization
Sicheng Liu, Lintao Wang, Xiaogang Zhu +3
Extreme Multimodal Summarization with Multimodal Output (XMSMO) becomes an attractive summarization approach by integrating various types of information to create extremely concise…
cs.CV2023
Multi-Scale Control Signal-Aware Transformer for Motion Synthesis without Phase
Lintao Wang, Kun Hu, Lei Bai +3
Synthesizing controllable motion for a character using deep learning has been a promising approach due to its potential to learn a compact model without laborious feature engineeri…