2 papers
cs.CV2026
Scalable Evaluation of the Realism of Synthetic Environmental Augmentations in Images
Damian J. Ruck, Paul Vautravers, Oliver Chalkley +1
Evaluation of AI systems often requires synthetic test cases, particularly for rare or safety-critical conditions that are difficult to observe in operational data. Generative AI o…
cs.CL2025
Text2VLM: Adapting Text-Only Datasets to Evaluate Alignment Training in Visual Language Models
Gabriel Downer, Sean Craven, Damian Ruck +1
The increasing integration of Visual Language Models (VLMs) into AI systems necessitates robust model alignment, especially when handling multimodal content that combines text and…