ملف الباحث
Xiaotong Ji
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
On Almost Surely Safe Alignment of Large Language Models at Inference-Time
2025 · arXiv (Cornell University)
We introduce a novel inference-time alignment approach for LLMs that aims to generate safe responses almost surely, i.e., with probability approaching one. Our approach models the generation of safe responses as a constrained Markov Decision …