Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
Hammer: Robust Function-Calling for On-Device Language Models via Function Masking
Qiqiang Lin, Muning Wen, Qiuying Peng +8
Large language models have demonstrated impressive value in performing as autonomous agents when equipped with external tools and API calls. Nonetheless, effectively harnessing the…
cs.LG2024
Entropy-Regularized Token-Level Policy Optimization for Language Agent Reinforcement
Muning Wen, Junwei Liao, Cheng Deng +3
Large Language Models (LLMs) have shown promise as intelligent agents in interactive decision-making tasks. Traditional approaches often depend on meticulously designed prompts, hi…