Shieldstral is a 3B open-weights multimodal safety classifier that outperforms larger models by up to 7x by framing content moderation as a policy-adaptive question-answering task. It accepts plain-language policies at inference time, unifying text and image safety evaluation without retraining, and is released under Apache 2.0, providing calibrated safety scores while operating efficiently on a single 16GB NVIDIA GPU.
mistral.ai
5 min
10h ago
Shieldstral is a 3B open-weights multimodal safety classifier that outperforms larger models by up to 7x by framing content moderation as a policy-adaptive question-answering task. It accepts plain-language policies at inference time, unifying text and image safety evaluation without retraining, and is released under Apache 2.0, providing calibrated safety scores while operating efficiently on a single 16GB NVIDIA GPU.
mistral.ai
5 min
10h ago
Shieldstral is a 3B open-weights multimodal safety classifier that outperforms larger models by up to 7x by framing content moderation as a policy-adaptive question-answering task. It accepts plain-language policies at inference time, unifying text and image safety evaluation without retraining, and is released under Apache 2.0, providing calibrated safety scores while operating efficiently on a single 16GB NVIDIA GPU.
mistral.ai
5 min
10h ago
No more articles to load