Researcher profile

Shalabh Bhatnagar

2 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. Bounds for off-policy prediction in reinforcement learning

    2017

    In this paper, we provide for the first time, error bounds for the off-policy prediction in reinforcement learning. The primary objective in off-policy prediction is to estimate the value function of a given target policy …

  2. Truncated Cauchy random perturbations for smoothed functional-based stochastic optimization

    2024 · Automatica

    In this paper, we present a stochastic gradient algorithm for minimizing a smooth objective function that is an expectation over noisy cost samples and only the latter are observed for any given parameter. Our algorithm …