Shieldstral
AssessTools
A 3B open-weights multimodal safety classifier for policy-adaptive moderation.
Why it's here
Placed in Assess: 1 article(s) of evidence from 1 source(s), led by model releases, with 1 in the last 30 days. Confidence 24%. Low accumulated evidence, so it defaults conservatively pending more signal.
Evidence (1)
- 7Hacker News·8/4/2026model_releaseMistral releases Shieldstral, a 3B multimodal safety model
Mistral has released Shieldstral, a 3B open-weights safety classifier designed for text and image moderation. It treats moderation as a policy-driven question-answering task, allowing plain-language policies to be supplied at inference time without retraining, and is released under Apache 2.0.