3 papers
cs.LG2026
BCPPO: Bachelier-Inspired Constrained Proximal Policy Optimization for Tail-Risk-Aware Safe Reinforcement Learning
Dongsheng Hou, Yanqiao Chen, Yuhan Rui
Expected-cost constraints can still permit rare, high-cost events. Monte Carlo conditional value at risk (CVaR) gradients can be noisy at high confidence, whereas critics that mode…
cs.AI2026
Shapley Context Pruning: A Cooperative Game Perspective for Context Reranking and Pruning
Yanqiao Chen, Dongsheng Hou, Yuhan Rui +2
Context reranking and pruning have become essential for improving the efficiency of modern Retrieval-Augmented Generation (RAG) systems, yet an interpretable and unified framework…
cs.CV2026
From Spatial to Spectral: An Efficient, Frequency-Guided Feature Representation Learner for Small Object Detection
Yuhan Rui, Shihan Qiao, Yibin Lou +7
Efficient small object detection is bottlenecked by the inherent feature scarcity of tiny targets, which is further aggravated by operations of spatial-domain detectors that indisc…