From the 1 of 8 linked papers with an AI index.
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
Steerable Pluralism: Pluralistic Alignment via Few-Shot Comparative Regression
Jadie Adams, Brian Hu, Emily Veenhuis +5
Large language models (LLMs) are currently aligned using techniques such as reinforcement learning from human feedback (RLHF). However, these methods use scalar rewards that can on…
cs.CL2025
ALIGN: Prompt-based Attribute Alignment for Reliable, Responsible, and Personalized LLM-based Decision-Making
Bharadwaj Ravichandran, David Joy, Paul Elliott +6
Large language models (LLMs) are increasingly being used as decision aids. However, users have diverse values and preferences that can affect their decision-making, which requires…
cs.CL2024
Defending Against Social Engineering Attacks in the Age of LLMs
Lin Ai, Tharindu Kumarage, Amrita Bhattacharjee +12
The proliferation of Large Language Models (LLMs) poses challenges in detecting and mitigating digital deception, as these models can emulate human conversational patterns and faci…