Mistral Releases Shieldstral, a 3B Open-Weight Model That Moderates Text and Images Against Plain-Language Policies at Inference Time
Mistral AI open-sourced Shieldstral, a 3B-parameter safety classifier that judges content against natural-language policies without retraining, released as the first project from NVIDIA's Open Secure AI Alliance.