7 citations · 19 across the 10 of their papers we have counts for
9 papers
Missingness-resilient Video-enhanced Multimodal Disfluency Detection
Payal Mohapatra, Shamika Likhite, Subrata Biswas +2
Most existing speech disfluency detection techniques only rely upon acoustic data. In this work, we present a practical multimodal disfluency detection approach that leverages avai…
Role Prompting Guided Domain Adaptation with General Capability Preserve for Large Language Models
Rui Wang, Fei Mi, Yi Chen +5
The growing interest in Large Language Models (LLMs) for specialized applications has revealed a significant challenge: when tailored to specific domains, LLMs tend to experience c…
Brain Functional Connectivity under Teleoperation Latency: a fNIRS Study
Yang Ye, Tianyu Zhou, Qi Zhu +2
Objective: This study aims to understand the cognitive impact of latency in teleoperation and the related mitigation methods, using functional Near-Infrared Spectroscopy (fNIRS) to…
Effect of Attention and Self-Supervised Speech Embeddings on Non-Semantic Speech Tasks
Payal Mohapatra, Akash Pandey, Yueyuan Sui +1
Human emotion understanding is pivotal in making conversational technology mainstream. We view speech emotion understanding as a perception task which is a more realistic setting.…
Explaining and Adapting Graph Conditional Shift
Qi Zhu, Yizhu Jiao, Natalia Ponomareva +2
Graph Neural Networks (GNNs) have shown remarkable performance on graph-structured data. However, recent empirical studies suggest that GNNs are very susceptible to distribution sh…
Collaborative Multi-Agent Video Fast-Forwarding
Shuyue Lan, Zhilu Wang, Ermin Wei +2
Multi-agent applications have recently gained significant popularity. In many computer vision tasks, a network of agents, such as a team of robots with cameras, could work collabor…