ملف الباحث
J. Nicholas Betley
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Connecting the Dots: LLMs can Infer and Verbalize Latent Structure from Disparate Training Data
2024 · arXiv (Cornell University)
One way to address safety risks from large language models (LLMs) is to censor dangerous knowledge from their training data. While this removes the explicit information, implicit information can remain scattered across various training documents. …