2 papers
cs.AI2026
LifeEval: A Multimodal Benchmark for Assistive AI in Egocentric Daily Life Tasks
Hengjian Gao, Kaiwei Zhang, Shibo Wang +8
The rapid progress of Multimodal Large Language Models (MLLMs) marks a significant step toward artificial general intelligence, offering great potential for augmenting human capabi…
eess.AS2025
Fast Word Error Rate Estimation Using Self-Supervised Representations for Speech and Text
Chanho Park, Chengsong Lu, Mingjie Chen +1
Word error rate (WER) estimation aims to evaluate the quality of an automatic speech recognition (ASR) system's output without requiring ground-truth labels. This task has gained i…