Researcher profile

Sedrick Scott Keh

2 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. A Critical Evaluation of AI Feedback for Aligning Large Language Models

    2024 · arXiv (Cornell University)

    Reinforcement learning with AI feedback (RLAIF) is a popular paradigm for improving the instruction-following abilities of powerful pre-trained language models. RLAIF first performs supervised fine-tuning (SFT) using demonstrations from a teacher model and then further …

  2. Asking More Informative Questions for Grounded Retrieval

    2024

    When a model is trying to gather information in an interactive setting, it benefits from asking informative questions.However, in the case of a grounded multi-turn image identification task, previous studies have been constrained to polar …