3 papers
cs.CL2026
A Content-Based Framework for Cybersecurity Refusal Decisions in Large Language Models
Noa Linder, Meirav Segal, Omer Antverg +5
Large language models and LLM-based agents are increasingly used for cybersecurity tasks that are inherently dual-use. Existing approaches to refusal, spanning academic policy fram…
cs.LG2025
Leveraging LLM Inconsistency to Boost Pass@k Performance
Uri Dalal, Meirav Segal, Zvika Ben-Haim +2
Large language models (LLMs) achieve impressive abilities in numerous domains, but exhibit inconsistent performance in response to minor input changes. Rather than view this as a d…
cs.LG2025
What Makes an Evaluation Useful? Common Pitfalls and Best Practices
Gil Gekker, Meirav Segal, Dan Lahav +1
Following the rapid increase in Artificial Intelligence (AI) capabilities in recent years, the AI community has voiced concerns regarding possible safety risks. To support decision…