Scale AI
@scale_AI
AI safety doesn’t translate one-to-one.
We partnered with the Korea AI Safety Institute to develop ROK-FORTRESS, our first benchmark together, testing how language and geopolitical context affect AI safety.
Across 14 frontier models, we found that translation alone can miss meaningful differences in model behavior. https://t.co/VnHnMNvWJ2