9 papers
Unified Agent: Managing Interactions across Devices
Xinshuang Liu, Runfa Blark Li, Shaoxiu Wei +2
As capabilities rapidly increase, AI agents can move from running inside one app to acting across a user's devices over time. Yet existing agent systems still fall short in this sc…
CoDoL: Conditional Domain Prompt Learning for Out-of-Distribution Generalization
Min Zhang, Yuyin Wang, Zhongxiang Dai +4
Recent advances in pre-training vision-language models (VLMs), e.g., contrastive language-image pre-training (CLIP) methods, have shown great potential in learning out-of-distribut…
MetaForge: A Self-Evolving Multimodal Agent that Retrieves, Adapts, and Forges Tools On Demand
Shouang Wei, Houcheng Min, Xinpeng Dong +8
Multimodal agents have achieved notable progress on complex reasoning tasks through tool use, yet remain limited by two issues: statically predefined tool inventories fail to gener…
UCO: A Multi-Turn Interactive Reinforcement Learning Method for Adaptive Teaching with Large Language Models
Shouang Wei, Min Zhang, Xin Lin +3
Large language models (LLMs) are shifting from answer providers to intelligent tutors in educational settings, yet current supervised fine-tuning methods only learn surface teachin…
SMRC: Aligning Large Language Models with Student Reasoning for Mathematical Error Correction
Biaojie Zeng, Min Zhang, Juan Zhou +3
Large language models (LLMs) often make reasoning errors when solving mathematical problems, and how to automatically detect and correct these errors has become an important resear…
Tighter Truncated Rectangular Prism Approximation for RNN Robustness Verification
Xingqi Lin, Liangyu Chen, Min Wu +2
Robustness verification is a promising technique for rigorously proving Recurrent Neural Networks (RNNs) robustly. A key challenge is to over-approximate the nonlinear activation f…