Koray Kavukcuoglu
4 papers in the PaperMetrix corpus
Papers by this author
-
Reinforcement Learning with Unsupervised Auxiliary Tasks
2016 · arXiv (Cornell University)
Deep reinforcement learning agents have achieved state-of-the-art results by directly maximising cumulative reward. However, environments contain a much wider variety of possible training signals. In this paper, we introduce an agent that also maximises many …
-
Neural Machine Translation in Linear Time
2016 · arXiv (Cornell University)
We present a novel neural network for processing sequences. The ByteNet is a one-dimensional convolutional neural network that is composed of two parts, one to encode the source sequence and the other to decode the …
-
Interaction Networks for Learning about Objects, Relations and Physics
2016 · arXiv (Cornell University)
Reasoning about objects, relations, and physics is central to human intelligence, and a key goal of artificial intelligence. Here we introduce the interaction network, a model which can reason about how objects in complex systems …
-
Scaling Language Models: Methods, Analysis & Insights from Training Gopher
2021 · arXiv (Cornell University)
Language modelling provides a step towards intelligent communication systems by harnessing large repositories of written human knowledge to better predict and understand the world. In this paper, we present an analysis of Transformer-based language model …