Researcher profile
Zhang, Ivan
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Mitigating harm in language models with conditional-likelihood filtration
2021 · arXiv (Cornell University)
Language models trained on large-scale unfiltered datasets curated from the open web acquire systemic biases, prejudices, and harmful views from their training data. We present a methodology for programmatically identifying and removing harmful text from …