2 papers
cs.AI2026
A Pipeline for Generating Longitudinal Synthetic Clinical Notes Using Large Language Models
William Poulett, Alice Waterhouse, Ben Wallace +4
Synthetic data is increasingly used to enable the development and evaluation of AI systems in domains where access to real-world data is restricted. In healthcare, clinical documen…
cs.CL2026
EvalSense: A Framework for Domain-Specific LLM (Meta-)Evaluation
Adam Dejl, Jonathan Pearson
Robust and comprehensive evaluation of large language models (LLMs) is essential for identifying effective LLM system configurations and mitigating risks associated with deploying…