5 papers
Agent-GWO: Collaborative Agents for Dynamic Prompt Optimization in Large Language Models
Xudong Wang, Chaoning Zhang, Chenghao Li +10
Large Language Models (LLMs) have demonstrated strong capabilities in complex reasoning tasks, while recent prompting strategies such as Chain-of-Thought (CoT) have further elevate…
Experience Transfer for Multimodal LLM Agents in Minecraft Game
Chenghao Li, Jun Liu, Songbo Zhang +7
Multimodal LLM agents operating in complex game environments must continually reuse past experience to solve new tasks efficiently. In this work, we propose Echo, a transfer-orient…
Rethinking Input Domains in Physics-Informed Neural Networks via Geometric Compactification Mappings
Zhenzhen Huang, Haoyu Bian, Jiaquan Zhang +6
Several complex physical systems are governed by multi-scale partial differential equations (PDEs) that exhibit both smooth low-frequency components and localized high-frequency st…
Text summarization via global structure awareness
Jiaquan Zhang, Chaoning Zhang, Shuxu Chen +9
Text summarization is a fundamental task in natural language processing (NLP), and the information explosion has made long-document processing increasingly demanding, making summar…
LLaVA-FA: Learning Fourier Approximation for Compressing Large Multimodal Models
Pengcheng Zheng, Chaoning Zhang, Jiarong Mo +8
Large multimodal models (LMMs) have achieved impressive performance on various vision-language tasks, but their substantial computational and memory costs hinder their practical de…