Mistral released Shieldstral 1.0, a 3-billion-parameter open-weights safety classifier (Apache 2.0) that judges text and images against a moderation policy written in plain language at inference time, returning calibrated probability scores instead of fixed categories. Mistral says it matches or beats open guard models up to seven times its size and runs on a single 16 GB GPU.