2 citations · 2 across the 3 of their papers we have counts for
3 papers
See, Hear, and Understand: Benchmarking Audiovisual Human Speech Understanding in Multimodal Large Language Models
Le Thien Phuc Nguyen, Zhuoran Yu, Samuel Low Yu Hang +8
Multimodal large language models (MLLMs) are expected to jointly interpret vision, audio, and language, yet existing video benchmarks rarely assess fine-grained reasoning about hum…
Understanding Human Daily Experience Through Continuous Sensing: ETRI Lifelog Dataset 2024
Se Won Oh, Hyuntae Jeong, Seungeun Chung +4
Improving human health and well-being requires an accurate and effective understanding of an individual's physical and mental state throughout daily life. To support this goal, we…
Human Understanding AI Paper Challenge 2024 -- Dataset Design
Se Won Oh, Hyuntae Jeong, Jeong Mook Lim +2
In 2024, we will hold a research paper competition (the third Human Understanding AI Paper Challenge) for the research and development of artificial intelligence technologies to un…