most citedo3-mini vs DeepSeek-R1: Which One is Safer?

3 citations · 6 across the 5 of their papers we have counts for

collaborators

5 papers

cs.SE20253 cited

o3-mini vs DeepSeek-R1: Which One is Safer?

Aitor Arrieta, Miriam Ugarte, Pablo Valle +2

The irruption of DeepSeek-R1 constitutes a turning point for the AI industry in general and the LLMs in particular. Its capabilities have demonstrated outstanding performance in se…

cs.SE20251 cited

Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation

Aitor Arrieta, Miriam Ugarte, Pablo Valle +2

Large Language Models (LLMs) have become an integral part of our daily lives. However, they impose certain risks, including those that can harm individuals' privacy, perpetuate bia…

cs.SE2025

ASTRAL: Automated Safety Testing of Large Language Models

Miriam Ugarte, Pablo Valle, José Antonio Parejo +2

Large Language Models (LLMs) have recently gained attention due to their ability to understand and generate sophisticated human-like content. However, ensuring their safety is para…

cs.SE20241 cited

Racing the Market: An Industry Support Analysis for Pricing-Driven DevOps in SaaS

Alejandro Garcia-Fernández, José Antonio Parejo, Francisco Javier Cavero +1

The SaaS paradigm has popularized the usage of pricings, allowing providers to offer customers a wide range of subscription possibilities. This creates a vast configuration space f…

cs.SE20241 cited

Quantum software experiments: A reporting and laboratory package structure guidelines

Enrique Moguel, José Antonio Parejo, Antonio Ruiz-Cortés +2

Background. In the realm of software engineering, there are widely accepted guidelines for reporting and creating laboratory packages. Unfortunately, the landscape differs consider…