31 citations · 31 across the 2 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2020
Leveraging the Variance of Return Sequences for Exploration Policy
Zerong Xi, Gita Sukthankar
This paper introduces a method for constructing an upper bound for exploration policy using either the weighted variance of return sequences or the weighted temporal difference (TD…
cs.LG2016
metricDTW: local distance metric learning in Dynamic Time Warping
Jiaping Zhao, Zerong Xi, Laurent Itti
We propose to learn multiple local Mahalanobis distance metrics to perform k-nearest neighbor (kNN) classification of temporal sequences. Temporal sequences are first aligned by dy…