8 citations · 9 across the 3 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2025
SELMA: A Speech-Enabled Language Model for Virtual Assistant Interactions
Dominik Wagner, Alexander Churchill, Siddharth Sigtia +1
In this work, we present and evaluate SELMA, a Speech-Enabled Language Model for virtual Assistant interactions that integrates audio and text as inputs to a Large Language Model (…
cs.SD2023★ 1 cited
Multimodal Data and Resource Efficient Device-Directed Speech Detection with Large Foundation Models
Dominik Wagner, Alexander Churchill, Siddharth Sigtia +4
Interactions with virtual assistants typically start with a trigger phrase followed by a command. In this work, we explore the possibility of making these interactions more natural…