Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
When Model Merging Rivals Joint Multi-Task Reinforcement Learning: A Task-Vector Geometry Analysis
S. Aaron McClendon
Model merging is promoted as a substitute for joint multi-task training, yet in the reinforcement-learning setting this substitution is essentially never tested against the baselin…
cs.LG2025
Reinforcement Learning for Machine Learning Model Deployment: Evaluating Multi-Armed Bandits in ML Ops Environments
S. Aaron McClendon, Vishaal Venkatesh, Juan Morinelli
In modern ML Ops environments, model deployment is a critical process that traditionally relies on static heuristics such as validation error comparisons and A/B testing. However,…