Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
WorldLines: Benchmarking and Modeling Long-Horizon Stateful Embodied Agents
Yehang Zhang, Jianchong Su, Haojian Huang +7
To assist humans over extended periods in real homes, embodied agents must remember user routines, world states, and past interactions. Existing long-term memory benchmarks mainly…
cs.AI2026
Code-in-the-Loop Forensics: Agentic Tool Use for Image Forgery Detection
Fanrui Zhang, Qiang Zhang, Sizhuo Zhou +10
Existing image forgery detection (IFD) methods either exploit low-level, semantics-agnostic artifacts or rely on multimodal large language models (MLLMs) with high-level semantic k…
cs.AI2026
Closing the Expression Gap in LLM Instructions via Socratic Questioning
Jianwen Sun, Yukang Feng, Yifan Chang +6
A fundamental bottleneck in human-AI collaboration is the ``intention expression gap," the difficulty for humans to effectively convey complex, high-dimensional thoughts to AI. Thi…