Researcher profile
Johannes Treutlein
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Connecting the Dots: LLMs can Infer and Verbalize Latent Structure from Disparate Training Data
2024 · arXiv (Cornell University)
One way to address safety risks from large language models (LLMs) is to censor dangerous knowledge from their training data. While this removes the explicit information, implicit information can remain scattered across various training documents. …