Researcher profile
Shalabh Bhatnagar
2 papers in the PaperMetrix corpus
Publications
Papers by this author
-
Bounds for off-policy prediction in reinforcement learning
2017
In this paper, we provide for the first time, error bounds for the off-policy prediction in reinforcement learning. The primary objective in off-policy prediction is to estimate the value function of a given target policy …
-
Truncated Cauchy random perturbations for smoothed functional-based stochastic optimization
2024 · Automatica
In this paper, we present a stochastic gradient algorithm for minimizing a smooth objective function that is an expectation over noisy cost samples and only the latter are observed for any given parameter. Our algorithm …