2 papers
cs.AI2026
Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games
Yidong He, Yutao Lai, Pengxu Yang +4
While Large Language Models (LLMs) excel in certain reasoning tasks, they struggle in multi-agent games where the final outcome depends on the joint strategies of all agents. In mu…
cs.SE2024
Is Your AI-Generated Code Really Safe? Evaluating Large Language Models on Secure Code Generation with CodeSecEval
Jiexin Wang, Xitong Luo, Liuwen Cao +5
Large language models (LLMs) have brought significant advancements to code generation and code repair, benefiting both novice and experienced developers. However, their training us…