Mistral's open model Shieldstral matches much larger safety models at a fraction of the size

Mistral's new 3B Shieldstral model checks AI inputs and outputs for safety violations using natural language yes-or-no questions instead of fixed categories. It matches models seven times its size in some benchmarks. Operators can set their own criteria at runtime rather than rely on a third party's category system, and the model can run locally. The article Mistral's open model Shieldstral…
This is a summary curated by AIFuture. Read the complete article at the original source:
Read the full story on The Decoder