2 papers
cs.LG2026
Adaptive Helpfulness-Harmlessness Alignment with Preference Vectors
Ren-Wei Liang, Chin-Ting Hsu, Chan-Hung Yu +6
Ensuring that large language models (LLMs) are both helpful and harmless is a critical challenge, as overly strict constraints can lead to excessive refusals, while permissive mode…
cs.LG2025
Synthesizing Programmatic Reinforcement Learning Policies with Large Language Model Guided Search
Max Liu, Chan-Hung Yu, Wei-Hsu Lee +3
Programmatic reinforcement learning (PRL) has been explored for representing policies through programs as a means to achieve interpretability and generalization. Despite promising…