Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
GTA-2: Benchmarking General Tool Agents from Atomic Tool-Use to Open-Ended Workflows
Jize Wang, Xuanxuan Liu, Yining Li +7
The development of general-purpose agents requires a shift from executing simple instructions to completing complex, real-world productivity workflows. However, current tool-use be…
cs.CL2025
EmbedGrad: Gradient-Based Prompt Optimization in Embedding Space for Large Language Models
Xiaoming Hou, Jiquan Zhang, Zibin Lin +2
Effectively adapting powerful pretrained foundation models to diverse tasks remains a key challenge in AI deployment. Current approaches primarily follow two paradigms:discrete opt…