3 papers
cs.CL2025
DPRM: A Dual Implicit Process Reward Model in Multi-Hop Question Answering
Xinyi Wang, Yiping Song, Zhiliang Tian +3
In multi-hop question answering (MHQA) tasks, Chain of Thought (CoT) improves the quality of generation by guiding large language models (LLMs) through multi-step reasoning, and Kn…
cs.AI2025
RLJP: Legal Judgment Prediction via First-Order Logic Rule-enhanced with Large Language Models
Yue Zhang, Zhiliang Tian, Shicheng Zhou +7
Legal Judgment Prediction (LJP) is a pivotal task in legal AI. Existing semantic-enhanced LJP models integrate judicial precedents and legal knowledge for high performance. But the…
cs.CL2025
DiffusionAttacker: Diffusion-Driven Prompt Manipulation for LLM Jailbreak
Hao Wang, Hao Li, Junda Zhu +4
Large Language Models (LLMs) are susceptible to generating harmful content when prompted with carefully crafted inputs, a vulnerability known as LLM jailbreaking. As LLMs become mo…