2 papers
cs.CR2026
Jailbreak Foundry: From Papers to Runnable Attacks for Reproducible Benchmarking
Zhicheng Fang, Jingjie Zheng, Chenxu Fu +1
Jailbreak techniques for large language models (LLMs) evolve faster than benchmarks, making robustness estimates stale and difficult to compare across papers due to drift in datase…
cs.AI2025
Clarifying Before Reasoning: A Coq Prover with Structural Context
Yanzhen Lu, Hanbin Yang, Xiaodie Wang +6
In this work, we investigate whether improving task clarity can enhance reasoning ability of large language models, focusing on theorem proving in Coq. We introduce a concept-level…