Scale AI Reports ROK-FORTRESS Findings on Multilingual AI Safety
Unite.AI
Read full postScale AI and the Korea AI Safety Institute released ROK-FORTRESS, a bilingual English-Korean adversarial safety benchmark evaluating 14 frontier AI models. The study found Korean-language prompts grounded in Korean contexts led to lower measured harm across models. The benchmark assesses responses in national security and public safety domains using calibrated LLM judges and expert rubrics.




