3 papers
cs.NI2026
FaulT-Bench: Towards Benchmarking Network Troubleshooting LLM Agents under Unreliable User Tickets
Kuan-Hao Tseng, Niruth Bogahawatta, Yasod Ginige +3
LLM-based agents are increasingly proposed for network fault diagnosis, but existing benchmarks evaluate them only on accurate tickets and always assume a fault is present, conditi…
cs.CR2025
AutoPentester: An LLM Agent-based Framework for Automated Pentesting
Yasod Ginige, Akila Niroshan, Sajal Jain +1
Penetration testing and vulnerability assessment are essential industry practices for safeguarding computer systems. As cyber threats grow in scale and complexity, the demand for p…
cs.CL2025
A Framework to Assess Multilingual Vulnerabilities of LLMs
Likai Tang, Niruth Bogahawatta, Yasod Ginige +4
Large Language Models (LLMs) are acquiring a wider range of capabilities, including understanding and responding in multiple languages. While they undergo safety training to preven…