Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Adaptive Multilevel Twisted Sequential Monte Carlo for Rare Events Estimation in Language Models
Zixuan Liu, Fangzheng Wu, Brian Summa +1
Rare unsafe behaviors in large language models can remain practically significant even when their probability is extremely small, particularly at deployment scales involving millio…
cs.LG2026
Robust General Utility for Reinforcement Learning
Zixuan Liu, Fangzheng Wu, Brian Summa +1
Reinforcement learning (RL) with general utility extends classic RL by optimizing an arbitrary utility functional of the policy-induced occupancy measure, thereby enabling a broade…
cs.LG2024
Rapid and Precise Topological Comparison with Merge Tree Neural Networks
Yu Qin, Brittany Terese Fasy, Carola Wenk +1
Merge trees are a valuable tool in the scientific visualization of scalar fields; however, current methods for merge tree comparisons are computationally expensive, primarily due t…