2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.LG2023
Asymptotically Unbiased Off-Policy Policy Evaluation when Reusing Old Data in Nonstationary Environments
Vincent Liu, Yash Chandak, Philip Thomas +1
In this work, we consider the off-policy policy evaluation problem for contextual bandits and finite horizon reinforcement learning in the nonstationary setting. Reusing old data i…
cs.NI2022★ 2 cited
FP4: Line-rate Greybox Fuzz Testing for P4 Switches
Nofel Yaseen, Liangcheng Yu, Caleb Stanford +2
Compared to fixed-function switches, the flexibility of programmable switches comes at a cost, as programmer mistakes frequently result in subtle bugs in the network data plane. In…