Showing cs.HCShow all
2 papers · 1 filter
cs.HC2026
UNIPO: Unified Interactive Visual Explanation for RL Fine-Tuning Policy Optimization
Aeree Cho, Alexander D. Greenhalgh, Jonathan Bodea +2
Reinforcement learning has emerged as a dominant technique for fine-tuning the behavior of large language models, with policy optimization (PO) algorithms such as GRPO, DAPO, and D…
cs.HC2023
HPCClusterScape: Increasing Transparency and Efficiency of Shared High-Performance Computing Clusters for Large-scale AI Models
Heungseok Park, Aeree Cho, Hyojun Jeon +5
The emergence of large-scale AI models, like GPT-4, has significantly impacted academia and industry, driving the demand for high-performance computing (HPC) to accelerate workload…