Skip to main content
AI

Mistral's Shieldstral: 3B open-weights model for multimodal moderation

Mistral's Shieldstral: 3B open-weights model for multimodal moderation Image: Primary
Mistral AI released Shieldstral, a 3 billion parameter open-weights model for multimodal content moderation, the company announced Monday. The model frames moderation as a policy-adaptive question-answering task, accepting plain-language policies at inference time to evaluate both text and images without retraining. Shieldstral matches or outperforms open guard models up to seven times its size across text safety, refusal detection, and multimodal benchmarks, according to the company. It returns a calibrated safety score from a single forward pass and runs on a single 16GB NVIDIA GPU. The model unifies heterogeneous safety datasets by converting them into a shared instruction-query-document format and uses contrastive policy pairs to teach discrimination rather than memorization. Mistral combined complementary checkpoints via SLERP merging to recover policy calibration and adaptability in one model. Shieldstral is released under the Apache 2.0 license and is available for download as part of the Open Secure AI Alliance with NVIDIA and other organizations.
Sources
In this story
Published by Tech & Business, a media brand covering technology and business. This story was sourced from mistral.ai and reviewed by the T&B editorial agent team.
Back to Newswire
Keep reading
Full wire
AI
AI

Moonshot AI's Kimi K3 now available on Amazon Bedrock

AWS has made Kimi K3 from Moonshot AI available on Amazon Bedrock, describing it as the first open-weight model to reach 2.8 trillion parameters. The model combines native vision capabilities with a 1-million-token context window...

AI Infrastructure
AI Infrastructure

AWS launches GPU-aware SageMaker HyperPod Inference Gateway

AWS announced general availability of the SageMaker HyperPod Inference Gateway, a Kubernetes-native routing addon that places inference requests using real-time GPU signals such as KV cache utilization, queue depth, LoRA adapter r...

AI Infrastructure
AI Infrastructure

GMI Cloud reportedly seeks $300 million chip loan for Thailand

GMI Cloud, an Nvidia partner, is seeking a $300 million loan to buy chips for a facility in Thailand, according to people familiar with the matter. The prospective financing would add to similar Asian deals supporting AI-compute i...