Researcher profile
Varun Chandrasekaran
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Privately Aligning Language Models with Reinforcement Learning
2023 · arXiv (Cornell University)
Positioned between pre-training and user deployment, aligning large language models (LLMs) through reinforcement learning (RL) has emerged as a prevailing strategy for training instruction following-models such as ChatGPT. In this work, we initiate the study …