Mistral's Shieldstral matches much larger safety models
Mistral releases Shieldstral, a 3B safety model matching much larger models with customizable natural language criteria.
Mistral's new 3B Shieldstral model checks AI inputs and outputs for safety violations using natural language yes-or-no questions instead of fixed categories. It matches models seven times its size in some benchmarks. Operators can set their own criteria at runtime rather than rely on a third party's category system, and the model can run locally.
The article Mistral's open model Shieldstral matches much larger safety models at a fraction of the size appeared first on The Decoder.