Power Seeking - PromptRiskDB

Record summary

A quick snapshot of what this page covers.

Techniques1Attack methods connected to this risk.

Mitigations0Defenses that may help with related attacks.

Domain7. AI System Safety, Failures, & LimitationsThe broad risk area this belongs to.

How this risk is described and categorized.

Domain7. AI System Safety, Failures, & Limitations

Subdomain7.1 > AI pursuing its own goals in conflict with human goals or values

Entity2 - AI

Intent1 - Intentional

Timing3 - Other

CategoryRogue AIs (Internal)

SubcategoryPower Seeking

Attack methods connected to this risk.

demonstrated

Methodtext_similarity_sqliteConfidence57%

Defenses that may help with related attacks.

No propagated mitigations. No defense is available through the connected attack methods.

Research source for this risk, when available.

Included resource

AuthorsHendrycks, Mazzeika & WoodsideYear2023TypePreprint

Original source

Open the public repository used for AI risk records and taxonomy fields.