Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Deferred Audio Pruning with Local Audio-Visual Dynamics for Omni-LLMs
Kyeongyoon Lee, Hongyeob Kim, Youngeun Kim +1
Omni-modal LLMs jointly process audio, video, and text, but long multimodal sequences incur substantial prefill and KV-cache costs. Existing omni-modal compression methods primaril…
cs.AI2026
SkillRet: A Large-Scale Benchmark for Skill Retrieval in LLM Agents
Hongcheol Cho, Ryangkyung Kang, Youngeun Kim
As LLM agents are increasingly deployed with large libraries of reusable skills, selecting the right skill for a user request has become a critical systems challenge. In small libr…