George Tucker
4 papers in the PaperMetrix corpus
Papers by this author
-
Model-Based Reinforcement Learning for Atari
2019 · arXiv (Cornell University)
Model-free reinforcement learning (RL) can be used to learn effective policies for complex tasks, such as Atari games, even from image observations. However, this typically requires very large amounts of interaction -- substantially more, in …
-
RL Unplugged: Benchmarks for Offline Reinforcement Learning.
2020 · arXiv (Cornell University)
Offline methods for reinforcement learning have a potential to help bridge the gap between reinforcement learning research and real-world applications. They make it possible to learn policies from offline datasets, thus overcoming concerns associated with …
-
Model Selection in Batch Policy Optimization
2021 · arXiv (Cornell University)
We study the problem of model selection in batch policy optimization: given a fixed, partial-feedback dataset and $M$ model classes, learn a policy with performance that is competitive with the policy derived from the best …
-
Gemini: A Family of Highly Capable Multimodal Models
2023 · arXiv (Cornell University)
This report introduces a new family of multimodal models, Gemini, that exhibit remarkable capabilities across image, audio, video, and text understanding. The Gemini family consists of Ultra, Pro, and Nano sizes, suitable for applications ranging …