Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
PITA: Preference-Guided Inference-Time Alignment for LLM Post-Training
Sarat Chandra Bobbili, Ujwal Dinesha, Dheeraj Narasimha +1
Inference-time alignment enables large language models (LLMs) to generate outputs aligned with end-user preferences without further training. Recent post-training methods achieve t…
cs.AI2025
Risk-Averse Finetuning of Large Language Models
Sapana Chaudhary, Ujwal Dinesha, Dileep Kalathil +1
We consider the challenge of mitigating the generation of negative or toxic content by the Large Language Models (LLMs) in response to certain prompts. We propose integrating risk-…