APromptRiskDBThreat intelligence atlas
AI Risk

Explainability

"Any action or procedure performed by a model with the intention of clarifying or detailing its internal functions."

AI Risk7. AI System Safety, Failures, & Limitations7.4 > Lack of transparency or interpretability2 - Post-deployment

Record summary

A quick snapshot of what this page covers.

Techniques0Attack methods connected to this risk.
Mitigations0Defenses that may help with related attacks.
Domain7. AI System Safety, Failures, & LimitationsThe broad risk area this belongs to.

Risk profile

How this risk is described and categorized.

Domain7. AI System Safety, Failures, & Limitations
Subdomain7.4 > Lack of transparency or interpretability
Entity2 - AI
Intent3 - Other
Timing2 - Post-deployment
CategoryExplainability
Subcategoryn/a

Suggested mitigations

Defenses that may help with related attacks.

No propagated mitigations. No defense is available through the connected attack methods.

Source

Research source for this risk, when available.