ملف الباحث

Denis Yarats

4 أوراق في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Automatic Data Augmentation for Generalization in Reinforcement Learning

    2021 · arXiv (Cornell University)

    Deep reinforcement learning (RL) agents often fail to generalize beyond their training environments. To alleviate this problem, recent work has proposed the use of data augmentation. However, different tasks tend to benefit from different types …

  2. CIC: Contrastive Intrinsic Control for Unsupervised Skill Discovery

    2022 · arXiv (Cornell University)

    We introduce Contrastive Intrinsic Control (CIC), an algorithm for unsupervised skill discovery that maximizes the mutual information between state-transitions and latent skill vectors. CIC utilizes contrastive learning between state-transitions and skills to learn behavior embeddings …

  3. Watch and Match: Supercharging Imitation with Regularized Optimal Transport

    2022 · arXiv (Cornell University)

    Imitation learning holds tremendous promise in learning policies efficiently for complex decision making problems. Current state-of-the-art algorithms often use inverse reinforcement learning (IRL), where given a set of expert demonstrations, an agent alternatively infers a …

  4. Convolutional Sequence to Sequence Learning

    2017 · arXiv (Cornell University)

    The prevalent approach to sequence to sequence learning maps an input sequence to a variable length output sequence via recurrent neural networks. We introduce an architecture based entirely on convolutional neural networks. Compared to recurrent …