ملف الباحث
Stephen Casper
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Black-Box Access is Insufficient for Rigorous AI Audits
2024
External audits of AI systems are increasingly recognized as a key mechanism for AI governance. The effectiveness of an audit, however, depends on the degree of access granted to auditors. Recent audits of state-of-the-art AI …
-
STACK: Adversarial Attacks on LLM Safeguard Pipelines
2025 · arXiv (Cornell University)
Frontier AI developers are relying on layers of safeguards to protect against catastrophic misuse of AI systems. Anthropic and OpenAI guard their latest Opus 4 model and GPT-5 models using such defense pipelines, and other …