3 papers
cs.CL2026
EVADE-Bench: Multimodal Benchmark for Evaluating and Enhancing Evasive Content Detection
Ancheng Xu, Zhihao Yang, Jingpeng Li +9
E-commerce platforms increasingly rely on Large Language Models (LLMs) and Vision Language Models (VLMs) to detect illicit or misleading product content. However, these models rema…
cs.HC2026
Making AI Drafts Count: A Quality Threshold in Audio Description Workflows
Lana Do, Shasta Ihorn, Charity M. Pitcher-Cooper +7
Audio description (AD) narrates visual elements in video for blind and low-vision audiences. Recent work has shown that giving novice describers an AI-generated draft to start from…
cs.AI2026
Structuring Reasoning for Complex Rules Beyond Flat Representations
Zhihao Yang, Ancheng Xu, Jingpeng Li +11
Large language models (LLMs) face significant challenges when processing complex rule systems, as they typically treat interdependent rules as unstructured textual data rather than…