3 papers
cs.CV2026
A Novel FACS-Aligned Anatomical Text Description Paradigm for Fine-Grained Facial Behavior Synthesis
Jiahe Wang, Cong Liang, Xuandong Huang +5
Facial behavior constitutes the primary medium of human nonverbal communication. Existing synthesis methods predominantly follow two paradigms: coarse emotion category labels or on…
cs.LG2024
Mixer is more than just a model
Qingfeng Ji, Yuxin Wang, Letong Sun
Recently, MLP structures have regained popularity, with MLP-Mixer standing out as a prominent example. In the field of computer vision, MLP-Mixer is noted for its ability to extrac…
cs.SD2024
ASM: Audio Spectrogram Mixer
Qingfeng Ji, Jicun Zhang, Yuxin Wang
Transformer structures have demonstrated outstanding skills in the deep learning space recently, significantly increasing the accuracy of models across a variety of domains. Resear…