3 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.CY2025★ 3 cited
Toward Quantitative Modeling of Cybersecurity Risks Due to AI Misuse
Steve Barrett, Malcolm Murray, Otter Quarks +17
Advanced AI systems offer substantial benefits but also introduce risks. In 2025, AI-enabled cyber offense has emerged as a concrete example. This technical report applies a quanti…
cs.LG2025
Leveraging LLM Inconsistency to Boost Pass@k Performance
Uri Dalal, Meirav Segal, Zvika Ben-Haim +2
Large language models (LLMs) achieve impressive abilities in numerous domains, but exhibit inconsistent performance in response to minor input changes. Rather than view this as a d…
cs.LG2025
What Makes an Evaluation Useful? Common Pitfalls and Best Practices
Gil Gekker, Meirav Segal, Dan Lahav +1
Following the rapid increase in Artificial Intelligence (AI) capabilities in recent years, the AI community has voiced concerns regarding possible safety risks. To support decision…