Researcher profile

Alan W. Black

13 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. Finding Function in Form: Compositional Character Models for Open Vocabulary Word Representation

    2015 · arXiv (Cornell University)

    Wang Ling, Chris Dyer, Alan W Black, Isabel Trancoso, Ramón Fermandez, Silvio Amir, Luís Marujo, Tiago Luís. Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing. 2015.

  2. Linguistic Markers of Influence in Informal Interactions

    2017

    There has been a long standing interest in understanding 'Social Influence' both in Social Sciences and in Computational Linguistics. In this paper, we present a novel approach to study and measure interpersonal influence in daily …

  3. A Survey of Code-switched Speech and Language Processing

    2019 · arXiv (Cornell University)

    Code-switching, the alternation of languages within a conversation or utterance, is a common communicative phenomenon that occurs in multilingual communities across the world. This survey reviews computational approaches for code-switched Speech and Natural Language Processing. …

  4. The Zero Resource Speech Challenge 2019: TTS Without T

    2019

    We present the Zero Resource Speech Challenge 2019, which proposes to build a\nspeech synthesizer without any text or phonetic labels: hence, TTS without T\n(text-to-speech without text). We provide raw audio for a target voice in …

  5. Task-Specific Pre-Training and Cross Lingual Transfer for Code-Switched Data

    2021 · arXiv (Cornell University)

    Using task-specific pre-training and leveraging cross-lingual transfer are two of the most popular ways to handle code-switched data. In this paper, we aim to compare the effects of both for the task of sentiment analysis. …

  6. DialoGraph: Incorporating Interpretable Strategy-Graph Networks into Negotiation Dialogues

    2021 · arXiv (Cornell University)

    To successfully negotiate a deal, it is not enough to communicate fluently: pragmatic planning of persuasive negotiation strategies is essential. While modern dialogue agents excel at generating fluent sentences, they still lack pragmatic grounding and …

  7. CodemixedNLP: An Extensible and Open NLP Toolkit for Code-Mixing

    2021

    The NLP community has witnessed steep progress in a variety of tasks across the realms of monolingual and multilingual language processing recently.

  8. Character-based Neural Machine Translation

    2015 · arXiv (Cornell University)

    We introduce a neural machine translation model that views the input and output sentences as sequences of characters rather than words. Since word-level information provides a crucial source of bias, our input model composes representations …

  9. Two/Too Simple Adaptations of Word2Vec for Syntax Problems

    2015

    Wang Ling, Chris Dyer, Alan W. Black, Isabel Trancoso. Proceedings of the 2015 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies. 2015.

  10. Polyglot Neural Language Models: A Case Study in Cross-Lingual Phonetic Representation Learning

    2016

    Yulia Tsvetkov, Sunayana Sitaram, Manaal Faruqui, Guillaume Lample, Patrick Littell, David Mortensen, Alan W Black, Lori Levin, Chris Dyer. Proceedings of the 2016 Conference of the North American Chapter of the Association for Computational Linguistics: …

  11. Style Transfer Through Back-Translation

    2018

    Style transfer is the task of rephrasing the text to contain specific stylistic properties without changing the intent or affect within the context. This paper introduces a new method for automatic style transfer. We first …

  12. A Dataset for Document Grounded Conversations

    2018

    This paper introduces a document grounded dataset for conversations. We define "Document Grounded Conversations" as conversations that are about the contents of a specified document. In this dataset the specified documents were Wikipedia articles about …

  13. Measuring Bias in Contextualized Word Representations

    2019

    Contextual word embeddings such as BERT have achieved state of the art performance in numerous NLP tasks. Since they are optimized to capture the statistical properties of training data, they tend to pick up on …