Researcher profile
Thanh Vinh Vo
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic
2025 · arXiv (Cornell University)
Hidden confounders that influence both states and actions can bias policy learning in reinforcement learning (RL), leading to suboptimal or non-generalizable behavior. Most RL algorithms ignore this issue, learning policies from observational trajectories based solely …