Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
Enhanced LLM Reasoning by Optimizing Reward Functions with Search-Driven Reinforcement Learning
Arash Ahmadi, Sarah Sharif, Yaser +1
Mathematical reasoning is a key benchmark for large language models. Reinforcement learning is a standard post-training mechanism for improving the reasoning capabilities of large…
cs.CL2026
Improving Heart-Focused Medical Question Answering in LLMs via Variance-Aware Rubric Rewards with GRPO
Arash Ahmadi, Parisa Masnadi, Parisa Masnadi Khiabani +4
Large Language Models (LLMs) have shown strong promise in healthcare applications. Yet deploying general-purpose models in real-world settings remains difficult due to data privacy…
cs.CL2025
Improving Aviation Safety Analysis: Automated HFACS Classification Using Reinforcement Learning with Group Relative Policy Optimization
Arash Ahmadi, Sarah Sharif, Yaser Banad
Analyzing the human factors behind aviation accidents is crucial for preventing future incidents, yet traditional methods using the Human Factors Analysis and Classification System…