ملف الباحث

Diyi Yang

8 أوراق في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Latent Hatred: A Benchmark for Understanding Implicit Hate Speech

    2021 · Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing

    Mai ElSherief, Caleb Ziems, David Muchlinski, Vaishnavi Anupindi, Jordyn Seybolt, Munmun De Choudhury, Diyi Yang. Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing. 2021.

  2. CoAnnotating: Uncertainty-Guided Work Allocation between Human and Large Language Models for Data Annotation

    2023 · arXiv (Cornell University)

    Annotated data plays a critical role in Natural Language Processing (NLP) in training models and evaluating their performance. Given recent developments in Large Language Models (LLMs), models such as ChatGPT demonstrate zero-shot capability on many …

  3. Mapping the Increasing Use of LLMs in Scientific Papers

    2024 · arXiv (Cornell University)

    Scientific publishing lays the foundation of science by disseminating research findings, fostering collaboration, encouraging reproducibility, and ensuring that scientific knowledge is accessible, verifiable, and built upon over time. Recently, there has been immense speculation about …

  4. Distilling an End-to-End Voice Assistant Without Instruction Training Data

    2024 · arXiv (Cornell University)

    Voice assistants, such as Siri and Google Assistant, typically model audio and text separately, resulting in lost speech information and increased complexity. Recent efforts to address this with end-to-end Speech Large Language Models (LLMs) trained …

  5. Incorporating Word Correlation Knowledge into Topic Modeling

    2015

    This paper studies how to incorporate the external word correlation knowledge to improve the coherence of topic modeling. Existing topic models assume words are generated independently and lack the mechanism to utilize the rich similarity …

  6. The GEM Benchmark: Natural Language Generation, its Evaluation and Metrics

    2021

    Sebastian Gehrmann, Tosin Adewumi, Karmanya Aggarwal, Pawan Sasanka Ammanamanchi, Anuoluwapo Aremu, Antoine Bosselut, Khyathi Raghavi Chandu, Miruna-Adriana Clinciu, Dipanjan Das, Kaustubh Dhole, Wanyu Du, Esin Durmus, Ondřej Dušek, Chris Chinenye Emezue, Varun Gangal, Cristina Garbacea, …

  7. Is ChatGPT a General-Purpose Natural Language Processing Task Solver?

    2023

    Spurred by advancements in scale, large language models (LLMs) have demonstrated the ability to perform a variety of natural language processing (NLP) tasks zero-shot—i.e., without adaptation on downstream data. Recently, the debut of ChatGPT has …

  8. Can Large Language Models Transform Computational Social Science?

    2023 · Computational Linguistics

    Abstract Large language models (LLMs) are capable of successfully performing many language processing tasks zero-shot (without training data). If zero-shot LLMs can also reliably classify and explain social phenomena like persuasiveness and political ideology, then …