Mistral Moderation
Overview
Product details compiled from public sources, each with a citation.
Matrix Coverage
Where this product defends, by asset class and NIST CSF function. The Coverage column shows whether each asset is Primary, Secondary, or Adjacent to what the product does. The table omits empty rows and columns.
| Asset class | Protect | Detect | Coverage | Source |
|---|---|---|---|---|
| AI Orchestration Tools | Secondary | 2 | ||
| Runtime AI Data | Primary | 2 |
Framework Relevance
These frameworks include controls relevant to the asset classes Mistral Moderation defends. This is an editorial inference from the AI Defense Matrix asset-level crossmap, not a statement that Mistral AI implements these controls or is certified against them.
Expand Collapse
| Framework | Asset class | Relevant controls |
|---|---|---|
| NIST IR 8596 | AI Orchestration Tools | Agents as deployed artifacts (orchestration view; see AI Agent Identities row for the principal view); system prompts and templates |
| Runtime AI Data | Prompts (runtime); inference data | |
| CSA AI Controls Matrix | AI Orchestration Tools | Application and Interface Security; Supply Chain Management |
| Runtime AI Data | Data Security and Privacy Lifecycle Management; Application and Interface Security | |
| ISO 42001 | AI Orchestration Tools | A.6 AI system life cycle; A.5 Assessing impacts of AI systems |
| Runtime AI Data | A.7 Data for AI systems; A.8 Information for interested parties | |
| Google SAIF | AI Orchestration Tools | Secure the AI supply chain; application and pipeline security; agent orchestration controls |
| Runtime AI Data | Expand AI red-teaming; runtime input and output safety; prompt defense | |
| SANS Critical AI Security Guidelines | AI Orchestration Tools | Secure Agentic Systems and AI Autonomy Controls (defined function scope; execution isolation; API and function-call gating); Limit Model Behavior (focused functionality; access controls outside the model) |
| Runtime AI Data | Model I/O Handling (sanitize, validate, and filter inputs and outputs; segregate user and system prompts; multilayered prompt-injection defense); Conventional Security Controls (protect augmentation and RAG data with vector-store access controls and validation); Data Minimization and Obfuscation (limit sensitive prompt content; context-window management); Limit Model Behavior (AI guardrails) | |
| MITRE ATLAS | AI Orchestration Tools | AML.T0051 LLM Prompt Injection; AML.T0054 LLM Jailbreak; AML.T0016 Obtain Capabilities (malicious plugins) |
| Runtime AI Data | AML.T0051 LLM Prompt Injection; AML.T0054 LLM Jailbreak; AML.T0056 Extract LLM System Prompt | |
| OWASP AI Exchange | AI Orchestration Tools | Development-time threats: agent framework supply chain; runtime threats: plugin abuse, prompt injection via tools |
| Runtime AI Data | Input threats: prompt injection, adversarial inputs, evasion; runtime threats: RAG poisoning, memory tampering | |
| OWASP LLM Top 10 | AI Orchestration Tools | LLM01 Prompt Injection; LLM05 Improper Output Handling; LLM07 System Prompt Leakage; LLM10 Unbounded Consumption |
| Runtime AI Data | LLM01 Prompt Injection; LLM02 Sensitive Information Disclosure; LLM08 Vector and Embedding Weaknesses; LLM05 Improper Output Handling | |
| OWASP Agentic Security Top 10 | AI Orchestration Tools | ASI01 Agent Goal Hijack; ASI02 Tool Misuse and Exploitation; ASI05 Unexpected Code Execution (RCE); ASI07 Insecure Inter-Agent Communication; ASI08 Cascading Failures; ASI10 Rogue Agents |
| Runtime AI Data | ASI06 Memory & Context Poisoning; ASI01 Agent Goal Hijack (via prompt injection in runtime inputs) |
Provenance
Last sourced 2026-06-10.
Expand Collapse
Sources
- Mistral AI documentation home
- Mistral AI moderation and guardrailing documentation
- “When a guardrail is triggered, the request is blocked and a 403 error is returned.”
- “Custom guardrails let you declare moderation rules directly in your API requests, without manually calling the Moderation API and implementing threshold logic in your application code.”
Changelog
-
Added to the catalog from the Mistral AI documentation.
Found an error? Corrections are welcome. Suggest an edit.