2 papers
cs.LG2026
Distributionally Robust Token Optimization in RLHF
Yeping Jin, Jiaming Hu, Ioannis Ch. Paschalidis
Large Language Models (LLMs) tend to respond correctly to prompts that align well with the data they were trained and fine-tuned on. Yet, small shifts in wording, format, or langua…
cs.LG2025
Distributionally Robust Learning in Survival Analysis
Yeping Jin, Lauren Wise, Ioannis Ch. Paschalidis
We introduce an innovative approach that incorporates a Distributionally Robust Learning (DRL) approach into Cox regression to enhance the robustness and accuracy of survival predi…