3 papers
cs.CL2025
The Aloe Family Recipe for Open and Specialized Healthcare LLMs
Dario Garcia-Gasulla, Jordi Bayarri-Planas, Ashwin Kumar Gururajan +10
Purpose: With advancements in Large Language Models (LLMs) for healthcare, the need arises for competitive open-source models to protect the public interest. This work contributes…
cs.CL2025
Efficient Safety Retrofitting Against Jailbreaking for LLMs
Dario Garcia-Gasulla, Adrian Tormos, Anna Arias-Duart +4
Direct Preference Optimization (DPO) is an efficient alignment technique that steers LLMs towards preferable outputs by training on preference data, bypassing the need for explicit…
cs.CV2024
Data Augmentation with Diffusion Models for Colon Polyp Localization on the Low Data Regime: How much real data is enough?
Adrian Tormos, Blanca Llauradó, Fernando Núñez +3
The scarcity of data in medical domains hinders the performance of Deep Learning models. Data augmentation techniques can alleviate that problem, but they usually rely on functiona…