8 papers
Hackers or Hallucinators? A Comprehensive Analysis of LLM-Based Automated Penetration Testing
Jiaren Peng, Zeqin Li, Chang You +17
The rapid advancement of Large Language Models (LLMs) has created new opportunities for Automated Penetration Testing (AutoPT), spawning numerous frameworks aimed at achieving end-…
From Retrieval to Reasoning: A Framework for Cyber Threat Intelligence NER with Explicit and Adaptive Instructions
Jiaren Peng, Hongda Sun, Xuan Tian +3
The automation of Cyber Threat Intelligence (CTI) relies heavily on Named Entity Recognition (NER) to extract critical entities from unstructured text. Currently, Large Language Mo…
MockLLM: A Multi-Agent Behavior Collaboration Framework for Online Job Seeking and Recruiting
Hongda Sun, Hongzhan Lin, Haiyu Yan +3
Online recruitment platforms have reshaped job-seeking and recruiting processes, driving increased demand for applications that enhance person-job matching. Traditional methods gen…
Enhancing Medical Dialogue Generation through Knowledge Refinement and Dynamic Prompt Adjustment
Hongda Sun, Jiaren Peng, Wenzhong Yang +3
Medical dialogue systems (MDS) have emerged as crucial online platforms for enabling multi-turn, context-aware conversations with patients. However, existing MDS often struggle to…
MobileSteward: Integrating Multiple App-Oriented Agents with Self-Evolution to Automate Cross-App Instructions
Yuxuan Liu, Hongda Sun, Wei Liu +3
Mobile phone agents can assist people in automating daily tasks on their phones, which have emerged as a pivotal research spotlight. However, existing procedure-oriented agents str…
BiDeV: Bilateral Defusing Verification for Complex Claim Fact-Checking
Yuxuan Liu, Hongda Sun, Wenya Guo +4
Complex claim fact-checking performs a crucial role in disinformation detection. However, existing fact-checking methods struggle with claim vagueness, specifically in effectively…