2 papers
cs.SE2024
Can Language Models Replace Programmers for Coding? REPOCOD Says 'Not Yet'
Shanchao Liang, Yiran Hu, Nan Jiang +1
Recently, a number of repository-level code generation benchmarks-such as CoderEval, DevEval, RepoEval, RepoBench, and LongCodeArena-have emerged to evaluate the capabilities of la…
cs.RO2024
SELP: Generating Safe and Efficient Task Plans for Robot Agents with Large Language Models
Yi Wu, Zikang Xiong, Yiran Hu +5
Despite significant advancements in large language models (LLMs) that enhance robot agents' understanding and execution of natural language (NL) commands, ensuring the agents adher…