Yizhe Zhang
13 papers in the PaperMetrix corpus
Papers by this author
-
Contextual Text Style Transfer
2020 · arXiv (Cornell University)
We introduce a new task, Contextual Text Style Transfer - translating a sentence into a desired style with its surrounding context taken into account. This brings two key challenges to existing style transfer approaches: ($i$) …
-
Dialogue Response Ranking Training with Large-Scale Human Feedback Data
2020
Existing open-domain dialog models are generally trained to minimize the perplexity of target human responses. However, some human replies are more engaging than others, spawning more followup interactions. Current conversational models are increasingly capable of …
-
Deconvolutional Latent-Variable Model for Text Sequence Matching
2018 · Proceedings of the AAAI Conference on Artificial Intelligence
A latent-variable model is introduced for text matching, inferring sentence representations by jointly optimizing generative and discriminative objectives. To alleviate typical optimization challenges in latent-variable models for text, we employ deconvolutional networks as the sequence …
-
Jointly Optimizing Diversity and Relevance in Neural Response Generation
2019
Xiang Gao, Sungjin Lee, Yizhe Zhang, Chris Brockett, Michel Galley, Jianfeng Gao, Bill Dolan. Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 …
-
Deconvolutional Latent-Variable Model for Text Sequence Matching
2017 · arXiv (Cornell University)
A latent-variable model is introduced for text matching, inferring sentence representations by jointly optimizing generative and discriminative objectives. To alleviate typical optimization challenges in latent-variable models for text, we employ deconvolutional networks as the sequence …
-
Joint Embedding of Words and Labels for Text Classification
2018
Guoyin Wang, Chunyuan Li, Wenlin Wang, Yizhe Zhang, Dinghan Shen, Xinyuan Zhang, Ricardo Henao, Lawrence Carin. Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2018.
-
Baseline Needs More Love: On Simple Word-Embedding-Based Models and Associated Pooling Mechanisms
2018
Dinghan Shen, Guoyin Wang, Wenlin Wang, Martin Renqiang Min, Qinliang Su, Yizhe Zhang, Chunyuan Li, Ricardo Henao, Lawrence Carin. Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). …
-
Domain Adaptive Text Style Transfer
2019
Dianqi Li, Yizhe Zhang, Zhe Gan, Yu Cheng, Chris Brockett, Bill Dolan, Ming-Ting Sun. Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language …
-
DIALOGPT : Large-Scale Generative Pre-training for Conversational Response Generation
2020
Yizhe Zhang, Siqi Sun, Michel Galley, Yen-Chun Chen, Chris Brockett, Xiang Gao, Jianfeng Gao, Jingjing Liu, Bill Dolan. Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics: System Demonstrations. 2020.
-
A Controllable Model of Grounded Response Generation
2021 · Proceedings of the AAAI Conference on Artificial Intelligence
Current end-to-end neural conversation models inherently lack the flexibility to impose semantic control in the response generation process, often resulting in uninteresting responses. Attempts to boost informativeness alone come at the expense of factual accuracy, …
-
DialoGPT: Large-Scale Generative Pre-training for Conversational Response Generation
2019 · arXiv (Cornell University)
We present a large, tunable neural conversational response generation model, DialoGPT (dialogue generative pre-trained transformer). Trained on 147M conversation-like exchanges extracted from Reddit comment chains over a period spanning from 2005 through 2017, DialoGPT extends …
-
Optimus: Organizing Sentences via Pre-trained Modeling of a Latent Space
2020
When trained effectively, the Variational Autoencoder (VAE) In this paper, we propose the first large-scale language VAE model OPTIMUS 1 . A universal latent embedding space for sentences is first pre-trained on large text corpus, …
-
What Makes Good In-Context Examples for GPT-3?
2022
Jiachang Liu, Dinghan Shen, Yizhe Zhang, Bill Dolan, Lawrence Carin, Weizhu Chen. Proceedings of Deep Learning Inside Out (DeeLIO 2022): The 3rd Workshop on Knowledge Extraction and Integration for Deep Learning Architectures. 2022.