4 papers
PolyWorkBench: Benchmarking LLM Agents for Cross-Lingual Long-Horizon Workflows
Hongliang Li, Yijin Liu, Zhiwei Zhang +5
While Large Language Model (LLM) agents excel at monolingual long-horizon planning and tool use, enterprise workflows inherently require processing multilingual resources across ex…
Beyond English: Uncovering the Multilingual Gap in Vision-Language-Action Models
Hanyang Chen, Hongliang Li, Jiarui Cao +6
Vision-Language-Action models have recently demonstrated promising capabilities in learning generalist robot policies from large-scale multimodal data. However, most existing VLA s…
Multilingual Collaborative Defense for Large Language Models
Hongliang Li, Jinan Xu, Gengping Cui +3
The robustness and security of large language models (LLMs) has become a prominent research area. One notable vulnerability is the ability to bypass LLM safeguards by translating h…
Towards Cost-Effective Reward Guided Text Generation
Ahmad Rashid, Ruotian Wu, Rongqi Fan +3
Reward-guided text generation (RGTG) has emerged as a viable alternative to offline reinforcement learning from human feedback (RLHF). RGTG methods can align baseline language mode…