2 papers
cs.CL2026
KoNeoBench: A Curated Evaluation Dataset for LLM Understanding of Korean Neologisms
Soha Lee, Soojin Lee, Heesung Yang +8
Large language models (LLMs) are typically evaluated on static benchmarks, even though natural language constantly evolves through newly emerging words and meanings. Existing Korea…
cs.AI2026
Can Language Models Understand mmWave Data? Benchmarking Large Language Models for mmWave Radar-Based Human Understanding
Jeongwan Shin, Jaehyeon Kim, Donguk Ko +1
Large language models (LLMs) have shown remarkable reasoning and generative capabilities, motivating their use as universal reasoning engines for perception. While modern approache…