Publications (15)
Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
Yuhui Sun, Xiyao Wang, Zixi Li +6
Copy, Right? A Testing Framework for Copyright Protection of Deep Learning Models
Jialuo Chen, Jingyi Wang, Tinglan Peng +6
Taming OpenClaw: Security Analysis and Mitigation of Autonomous LLM Agent Threats
Xinhao Deng, Yixiang Zhang, Jiaqing Wu +15
FAST: Boosting Uncertainty-based Test Prioritization Methods for Neural Networks via Feature Selection
Jialuo Chen, Jingyi Wang, Xiyue Zhang +4
S-Eval: Towards Automated and Comprehensive Safety Evaluation for Large Language Models
Xiaohan Yuan, Jinfeng Li, Dongxia Wang +9
The RPM3D project: 3D Kinematics for Remote Patient Monitoring
Alicia Fornés, Asma Bensalah, Cristina Carmona-Duarte +9
ASEval: Automated Trajectory-Level Security Testing for Autonomous Agents
Jianan Ma, Xiaohu Du, Ruixiao Lin +9
RobOT: Robustness-Oriented Testing for Deep Learning Systems
Jingyi Wang, Jialuo Chen, Youcheng Sun +4
Towards Stroke Patients' Upper-limb Automatic Motor Assessment Using Smartwatches
Asma Bensalah, Jialuo Chen, Alicia Fornés +3
MICCAI STS 2024 Challenge: Semi-Supervised Instance-Level Tooth Segmentation in Panoramic X-ray and CBCT Images
Yaqi Wang, Zhi Li, Chengyu Wu +19
Forward Regression via Gram-Schmidt Orthogonalization for Ultra-High Dimensional Linear Models
Jialuo Chen, Zhaoxing Gao, Yifan Jiang +1
MICCAI STSR 2025 Challenge: Semi-Supervised Teeth and Pulp Segmentation and CBCT-IOS Registration
Yaqi Wang, Zhi Li, Chengyu Wu +15
Safety Testing LLM Agents at Scale: From Risk Discovery to Evidence-Grounded Verification
Yunhao Feng, Ruixiao Lin, Ming Wen +12
Structured Analysis and Comparison of Alphabets in Historical Handwritten Ciphers
MartÃn Méndez, Pau Torras, Adrià Molina +3
Piggybacking on Perception: Stealthy Concurrent Audio Prompt Injections against Multimodal LLM Agents
Mingxiao Liu, Yitong Li, Haoren Zhao +6
The paper studies stealthy audio prompt injection attacks that hide malicious instructions within normal speech to hijack multimodal LLM agents, introduces a benchmark (AudioAgentS…