From the 1 of 70 linked papers with an AI index.
1 citations · 1 across the 18 of their papers we have counts for
5 papers · 1 filter
Kalman Filter Enhanced GRPO for Reinforcement Learning-Based Language Model Reasoning
Hu Wang, Congbo Ma, Ian Reid +1
The advantage function is a central concept in RL that helps reduce variance in policy gradient estimates. For language modeling, Group Relative Policy Optimization (GRPO) was prop…
Beyond Generative AI: World Models for Clinical Prediction, Counterfactuals, and Planning
Mohammad Areeb Qazi, Maryam Nadeem, Mohammad Yaqub
Healthcare requires AI that is predictive, reliable, and data-efficient. However, recent generative models lack physical foundation and temporal reasoning required for clinical dec…
Rethinking Weight-Averaged Model-merging
Hu Wang, Congbo Ma, Ibrahim Almakky +3
Model merging, particularly through weight averaging, has shown surprising effectiveness in saving computations and improving model performance without any additional training. How…
Forget-MI: Machine Unlearning for Forgetting Multimodal Information in Healthcare Settings
Shahad Hardan, Darya Taratynova, Abdelmajid Essofi +2
Privacy preservation in AI is crucial, especially in healthcare, where models rely on sensitive patient data. In the emerging field of machine unlearning, existing methodologies st…
SurvCORN: Survival Analysis with Conditional Ordinal Ranking Neural Network
Muhammad Ridzuan, Numan Saeed, Fadillah Adamsyah Maani +2
Survival analysis plays a crucial role in estimating the likelihood of future events for patients by modeling time-to-event data, particularly in healthcare settings where predicti…