2 papers
cs.AI2026
AgentSelect: Benchmark for Narrative Query-to-Agent Recommendation
Yunxiao Shi, Wujiang Xu, Tingwei Chen +7
LLM agents are rapidly becoming the practical interface for task automation, yet the ecosystem lacks a principled way to choose among an exploding space of deployable configuration…
cs.CL2026
M-QUEST -- Meme Question-Understanding Evaluation on Semantics and Toxicity
Stefano De Giorgis, Ting-Chih Chen, Filip Ilievski
Internet memes are a powerful form of online communication, yet their nature and reliance on commonsense knowledge make toxicity detection challenging. Identifying key features for…