3 citations · 6 across the 5 of their papers we have counts for
5 papers
o3-mini vs DeepSeek-R1: Which One is Safer?
Aitor Arrieta, Miriam Ugarte, Pablo Valle +2
The irruption of DeepSeek-R1 constitutes a turning point for the AI industry in general and the LLMs in particular. Its capabilities have demonstrated outstanding performance in se…
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation
Aitor Arrieta, Miriam Ugarte, Pablo Valle +2
Large Language Models (LLMs) have become an integral part of our daily lives. However, they impose certain risks, including those that can harm individuals' privacy, perpetuate bia…
ASTRAL: Automated Safety Testing of Large Language Models
Miriam Ugarte, Pablo Valle, José Antonio Parejo +2
Large Language Models (LLMs) have recently gained attention due to their ability to understand and generate sophisticated human-like content. However, ensuring their safety is para…
Racing the Market: An Industry Support Analysis for Pricing-Driven DevOps in SaaS
Alejandro Garcia-Fernández, José Antonio Parejo, Francisco Javier Cavero +1
The SaaS paradigm has popularized the usage of pricings, allowing providers to offer customers a wide range of subscription possibilities. This creates a vast configuration space f…
Quantum software experiments: A reporting and laboratory package structure guidelines
Enrique Moguel, José Antonio Parejo, Antonio Ruiz-Cortés +2
Background. In the realm of software engineering, there are widely accepted guidelines for reporting and creating laboratory packages. Unfortunately, the landscape differs consider…