ai safety 1alignment 1audit 1background manipulation 1benchmark dataset 1benchmarking 1detection 1evaluation 1image forensics 1large language models 1localization 1
From the 2 of 2 linked papers with an AI index.
2 papers
cs.CV2026
BG-REAL: A Public Real-Data Anchored Benchmark for Background Manipulation Detection and Localization
Bugra Alperen Uluirmak, Rifat Kurban
The paper introduces BG-REAL, a publicly available benchmark for detecting and localizing manipulations that occur in the background of images, built from real Open Images data and…
cs.AI2026
EvalSafetyGap: A Hybrid Survey and Conceptual Framework for LLM Evaluation-Safety Failures
BuÄra Alperen Uluırmak, Rifat Kurban
The paper surveys recent work on evaluating large language models (LLMs) for safety and introduces the EvalSafetyGap framework to compare evaluation and alignment failures, illustr…