12 papers
False Friends in the Shell: Unveiling the Emoticon Semantic Confusion in Large Language Models
Weipeng Jiang, Xiaoyu Zhang, Juan Zhai +3
Emoticons are widely used in digital communication to convey affective intent, yet their safety implications for Large Language Models (LLMs) remain largely unexplored. In this pap…
Weaponizing the Commons: A Taxonomy and Detection Framework of Abuse on GitHub
Yuli Cheng, Xiaoyu Zhang, Jiongchi Yu +3
GitHub plays a critical role in modern software supply chains, making its security an important research concern. Existing studies have primarily focused on CI/CD automation, colla…
Where Agent Frameworks Fall Short: Examining Functional Challenges and Usability Concerns
Xinxue Zhu, Jiacong Wu, Xiaoyu Zhang +6
Large language model (LLM) agents are increasingly built on agent frameworks that provide reusable abstractions for workflow orchestration, state management, tool integration, and…
CITADEL: Context Similarity Based Deep Learning Framework Bug Finding
Xiaoyu Zhang, Juan Zhai, Shiqing Ma +2
With the application of deep learning technology, tools of DL framework testing are in high demand. Existing DL framework testing tools have limited coverage of bug types. For exam…
MCP-RADAR: A Multi-Dimensional Benchmark for Evaluating Tool Use Capabilities in Large Language Models
Xuanqi Gao, Siyi Xie, Juan Zhai +2
As Large Language Models (LLMs) evolve from passive text generators to active reasoning agents capable of interacting with external tools, the Model Context Protocol (MCP) has emer…
Mitigating Stylistic Biases of Machine Translation Systems via Monolingual Corpora Only
Xuanqi Gao, Weipeng Jiang, Juan Zhai +4
The advent of neural machine translation (NMT) has revolutionized cross-lingual communication, yet preserving stylistic nuances remains a significant challenge. While existing appr…