Pablo Samuel Castro
3 papers in the PaperMetrix corpus
Papers by this author
-
Combining Learned Lyrical Structures and Vocabulary for Improved Lyric Generation
2018 · arXiv (Cornell University)
The use of language models for generating lyrics and poetry has received an increased interest in the last few years. They pose a unique challenge relative to standard natural language problems, as their ultimate purpose …
-
An Atari Model Zoo for Analyzing, Visualizing, and Comparing Deep Reinforcement Learning Agents
2019
Much human and computational effort has aimed to improve how deep reinforcement learning (DRL) algorithms perform on benchmarks such as the Atari Learning Environment. Comparatively less effort has focused on understanding what has been learned …
-
A functional mirror ascent view of policy gradient methods with function approximation.
2021 · arXiv (Cornell University)
We use functional mirror ascent to propose a general framework (referred to as FMA-PG) for designing policy gradient methods. The functional perspective distinguishes between a policy's functional representation (what are its sufficient statistics) and its …