3 papers
cs.CL2026
A Multi-Domain Red Teaming Framework for Safety, Robustness, and Fairness Evaluation of Medical Large Language Models
Andrei Marian Feier, Veysel Kocaman, Yigit Gul +6
Large language models (LLMs) are increasingly deployed across healthcare, yet existing benchmarks fail to capture model behavior under adversarial or ethically complex conditions c…
cs.CL2025
Can Zero-Shot Commercial APIs Deliver Regulatory-Grade Clinical Text DeIdentification?
Veysel Kocaman, Muhammed Santas, Yigit Gul +2
We evaluate the performance of four leading solutions for de-identification of unstructured medical text - Azure Health Data Services, AWS Comprehend Medical, OpenAI GPT-4o, and Jo…
cs.CL2025
Beyond Negation Detection: Comprehensive Assertion Detection Models for Clinical NLP
Veysel Kocaman, Yigit Gul, M. Aytug Kaya +4
Assertion status detection is a critical yet often overlooked component of clinical NLP, essential for accurately attributing extracted medical facts. Past studies have narrowly fo…