Yannis Flet-Berliac
3 papers in the PaperMetrix corpus
Papers by this author
-
Adversarially Guided Actor-Critic
2021 · arXiv (Cornell University)
Despite definite success in deep reinforcement learning problems,\nactor-critic algorithms are still confronted with sample inefficiency in\ncomplex environments, particularly in tasks where efficient exploration is a\nbottleneck. These methods consider a policy (the actor) and a value …
-
SAAC: Safe Reinforcement Learning as an Adversarial Game of Actor-Critics
2022 · arXiv (Cornell University)
Although Reinforcement Learning (RL) is effective for sequential decision-making problems under uncertainty, it still fails to thrive in real-world systems where risk or safety is a binding constraint. In this paper, we formulate the RL …
-
PASTA: Pretrained Action-State Transformer Agents
2023 · arXiv (Cornell University)
Self-supervised learning has brought about a revolutionary paradigm shift in various computing domains, including NLP, vision, and biology. Recent approaches involve pre-training transformer models on vast amounts of unlabeled data, serving as a starting point …