Mistral's Shieldstral: a 3B open-weights safety classifier that beats a 20B model

Mistral released Shieldstral, a 3B-parameter open-weights content-moderation classifier that reportedly outperforms models seven times its size by framing moderation as a plain-English yes/no questio…