Sylvain Lamprier
4 papers in the PaperMetrix corpus
Papers by this author
-
Using BERT and BART for Query Suggestion
2020 · HAL (Le Centre pour la Communication Scientifique Directe)
International audience
-
To Beam Or Not To Beam: That is a Question of Cooperation for Language GANs
2021 · arXiv (Cornell University)
Due to the discrete nature of words, language GANs require to be optimized from rewards provided by discriminator networks, via reinforcement learning methods. This is a much harder setting than for continuous tasks, which enjoy …
-
Offline Reinforcement Learning of High-Quality Behaviors Under Robust Style Alignment
2026 · HAL (Le Centre pour la Communication Scientifique Directe)
We study offline reinforcement learning of style-conditioned policies using explicit style supervision via subtrajectory labeling functions. In this setting, aligning style with high task performance is particularly challenging due to distribution shift and inherent conflicts …
-
Answers Unite! Unsupervised Metrics for Reinforced Summarization Models
2019
Thomas Scialom, Sylvain Lamprier, Benjamin Piwowarski, Jacopo Staiano. Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP). 2019.