2 papers
eess.SY2026
Information-Driven Active Perception for k-step Predictive Safety Monitoring
Sumukha Udupa, Jie Fu
This work studies the synthesis of active perception policies for predictive safety monitoring in partially observable stochastic systems. Operating under strict sensing and commun…
cs.LG2025
Thompson Sampling in Online RLHF with General Function Approximation
Songtao Feng, Jie Fu
Reinforcement learning from human feedback (RLHF) has achieved great empirical success in aligning large language models (LLMs) with human preference, and it is of great importance…