Scale AI on September 17, 2026 published findings from ROK-FORTRESS, a bilingual English–Korean adversarial safety benchmark developed jointly with the Korea AI Safety Institute, reporting that prompts written in Korean and grounded in Korean contexts were consistently associated with lower measured harm across nearly all 14 frontier models evaluated. Benchmark Design: The Transcreation Matrix Most multilingual safety benchmarks translate a fixed prompt into another language while keeping the…