4 papers
WuYuEval: A Multi-Level Benchmark for Large Language Models in Solid Waste Management
Yi Zhang, Hongyang Wang, Zheng Hao Leong +11
Large language models (LLMs) are increasingly used as technical assistants, but their competence in solid waste management (SWM) remains difficult to assess because existing benchm…
Auditing Cascading Risks in Multi-Agent Systems via Semantic-Geometric Co-evolution
Zixun Luo, Yuhang Fan, Hengyu Lin +2
Large Language model (LLM)-based Multi-Agent Systems (MAS) are prone to cascading risks, where early-stage interactions remain semantically fluent and policy-compliant, yet the und…
AIAuditTrack: A Framework for AI Security system
Zixun Luo, Yuhang Fan, Yufei Li +3
The rapid expansion of AI-driven applications powered by large language models has led to a surge in AI interaction data, raising urgent challenges in security, accountability, and…
Experimental Study on Time Series Analysis of Lower Limb Rehabilitation Exercise Data Driven by Novel Model Architecture and Large Models
Hengyu Lin
This study investigates the application of novel model architectures and large-scale foundational models in temporal series analysis of lower limb rehabilitation motion data, aimin…