Mistral AI Releases Shieldstral 1.0 3B: An Open-Weights Policy-Adaptive Multimodal Safety Classifier Matching Models 7× Its Size
MarkTechPost
Read Full Article at MarkTechPost →Ad Slot — In-Article (728x90)
Mistral AI has released Shieldstral 1. 0 3B, an open-weights, policy-adaptive multimodal safety classifier that frames content moderation as a single yes/no question instead of a fixed harm taxonomy.
Operators supply the policy as a plain-language query at inference time and get back a calibrated safety score from one forward pass — no retraining required to re-target the model. Built on Ministral-3-3B-Base-2512 with a Pixtral vision encoder and trained on roughly 54. 1M samples, it reports 84.
This is a summary. For the full story, read the original article at MarkTechPost.
Original source: MarkTechPost