4 papers
TRACE Bench: Task-driven Roleplay Agentic Checklist Evaluation
Jiahui Zhang, Ziwei Zhang, Yipeng Wang +7
Roleplay evaluation should do more than assign a single score: it should reveal which role requirements were tested, which failed, and which dialogue evidence supports the judgment…
P2DNav: Panorama-to-Downview Reasoning for Zero-shot Vision-and-Language Navigation
Kai Sheng, Liuyi Wang, Haojie Dai +5
Vision-and-language navigation (VLN) requires an embodied agent to ground natural-language instructions into executable navigation actions in unseen environments. Existing zero-sho…
Cutscene Agent: An LLM Agent Framework for Automated 3D Cutscene Generation
Lanshan He, Haozhou Pang, Qi Gan +12
Cutscenes are carefully choreographed cinematic sequences embedded in video games and interactive media, serving as the primary vehicle for narrative delivery, character developmen…
Edge-Cloud Collaborative Satellite Image Analysis for Efficient Man-Made Structure Recognition
Kaicheng Sheng, Junxiao Xue, Hui Zhang
The increasing availability of high-resolution satellite imagery has created immense opportunities for various applications. However, processing and analyzing such vast amounts of…