Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Diffusion Blend: Inference-Time Multi-Preference Alignment for Diffusion Models
Min Cheng, Fatemeh Doudi, Dileep Kalathil +2
Reinforcement learning (RL) algorithms have been used recently to align diffusion models with downstream objectives such as aesthetic quality and text-image consistency by fine-tun…
cs.AI2025
Risk-Averse Finetuning of Large Language Models
Sapana Chaudhary, Ujwal Dinesha, Dileep Kalathil +1
We consider the challenge of mitigating the generation of negative or toxic content by the Large Language Models (LLMs) in response to certain prompts. We propose integrating risk-…