2 papers
cs.CL2025
From General Reasoning to Domain Expertise: Uncovering the Limits of Generalization in Large Language Models
Dana Alsagheer, Yang Lu, Abdulrahman Kamal +7
Recent advancements in Large Language Models (LLMs) have demonstrated remarkable capabilities in various domains. However, effective decision-making relies heavily on strong reason…
cs.CY2025
Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability
Dana Alsagheer, Abdulrahman Kamal, Mohammad Kamal +1
Reinforcement Learning from Human Feedback (RLHF) is central in aligning large language models (LLMs) with human values and expectations. However, the process remains susceptible t…