Extraction attack risk for AI

Last updated: Feb 07, 2025

Robustness

Inference risks

Amplified by generative AI

Description

An attribute inference attack is used to detect whether certain sensitive features can be inferred about individuals who participated in training a model. These attacks occur when an adversary has some prior knowledge about the training data and uses that knowledge to infer the sensitive data.

Why is extraction attack a concern for foundation models?

With a successful extraction attack, the attacker can perform further adversarial attacks to gain valuable information such as sensitive personal information or intellectual property.

Parent topic: AI risk atlas

We provide examples covered by the press to help explain many of the foundation models' risks. Many of these events covered by the press are either still evolving or have been resolved, and referencing them can help the reader understand the potential risks and work towards mitigations. Highlighting these examples are for illustrative purposes only.

Was the topic helpful?

0/1000

DescriptionCopy link to section

Why is extraction attack a concern for foundation models?Copy link to section

Related RisksCopy link to section

Description

Why is extraction attack a concern for foundation models?

Related Risks