2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CV2024
An Efficient and Streaming Audio Visual Active Speaker Detection System
Arnav Kundu, Yanzi Jin, Mohammad Sekhavat +3
This paper delves into the challenging task of Active Speaker Detection (ASD), where the system needs to determine in real-time whether a person is speaking or not in a series of v…
cs.CL2024★ 2 cited
OpenELM: An Efficient Language Model Family with Open Training and Inference Framework
Sachin Mehta, Mohammad Hossein Sekhavat, Qingqing Cao +8
The reproducibility and transparency of large language models are crucial for advancing open research, ensuring the trustworthiness of results, and enabling investigations into dat…