Opacity (the black box problem)

Record summary

A quick snapshot of what this page covers.

Techniques0Attack methods connected to this risk.

Mitigations0Defenses that may help with related attacks.

Domain7. AI System Safety, Failures, & LimitationsThe broad risk area this belongs to.

Risk profile

How this risk is described and categorized.

"Opacity surrounding the technical, internal decision-making processes of generative AI models is popularly known as the “black box problem.”277 Generative AI models, most ubiquitously built on deep neural networks with hundreds of billions of internal connections,278 have become so complex that their internal decision-making processes are no longer traceable or interpretable to even the most advanced expert observers. This means that, while the inputs and outputs of a system can be observed, developers cannot explain in detail why specific inputs correspond to specific outputs."

Domain7. AI System Safety, Failures, & Limitations

Subdomain7.4 > Lack of transparency or interpretability

Entity3 - Other

Intent2 - Unintentional

Timing3 - Other

CategoryTechnical and operational risks

SubcategoryOpacity (the black box problem)

Related techniques

Attack methods connected to this risk.

No linked attack methods. No AI attack method is connected to this risk in the current data.

Suggested mitigations

Defenses that may help with related attacks.

No propagated mitigations. No defense is available through the connected attack methods.

Source

Research source for this risk, when available.

Included resource

Regulating under Uncertainty: Governance Options for Generative AI

AuthorsG'sellYear2024TypeReport

DOI10.2139/ssrn.4918704 URLhttps://papers.ssrn.com/sol3/papers.cfm?abstract_id=4918704

Original source

MIT AI Risk Repository

Open the public repository used for AI risk records and taxonomy fields.

Repositoryhttps://airisk.mit.edu/