Authors

Isaac Holeman

Publications 1

Preprint Open access

MLCommons Jailbreak Benchmark v1.0

Modern AI systems are designed to refuse hazardous requests. A jailbreak is a prompt crafted to bypass those safeguards and elicit outputs that the system would normally refuse to provide. The MLCommons Jailbreak Benchmark v1.0 provides an end-to-end methodology for evaluating the robustness of large language models to …

Co-authors