Mistral AI released Shieldstral, a 3 billion parameter open-weights multimodal safety classifier under Apache 2.0 that answers plain-language yes or no policy questions at inference time. Operators ask whether content promotes violence or is unsafe for minors, and the model returns a calibrated score from a single token across text, images, and prompt-response pairs without retraining. Mistral says it matches open guard models up to seven times larger on text safety, leads multimodal moderation tests, and runs on a single 16GB GPU. Caveat: vendor benchmarks need independent checks, and novel adversarial policies can still slip through.