58 citations · 80 across the 4 of their papers we have counts for
4 papers
Activating Visual Context and Commonsense Reasoning through Masked Prediction in VLMs
Jiaao Yu, Shenwei Li, Mingjie Han +4
Recent breakthroughs in reasoning models have markedly advanced the reasoning capabilities of large language models, particularly via training on tasks with verifiable rewards. Yet…
Meta-Learning via Classifier(-free) Diffusion Guidance
Elvis Nava, Seijin Kobayashi, Yifei Yin +2
We introduce meta-learning algorithms that perform zero-shot weight-space adaptation of neural network models to unseen tasks. Our methods repurpose the popular generative image sy…
VibEmoji: Exploring User-authoring Multi-modal Emoticons in Social Communication
Pengcheng An, Ziqi Zhou, Qing Liu +4
Emoticons are indispensable in online communications. With users' growing needs for more customized and expressive emoticons, recent messaging applications begin to support (limite…
Multi-Classifier Interactive Learning for Ambiguous Speech Emotion Recognition
Ying Zhou, Xuefeng Liang, Yu Gu +2
In recent years, speech emotion recognition technology is of great significance in industrial applications such as call centers, social robots and health care. The combination of s…