Skip to main content
OpenAI

Disrupting malicious uses of AI

Explore OpenAI’s latest case studies and reports on how we detect and disrupt malicious uses of AI.

Yellow grid-patterned card with purple text reading “Scam Spotting with ChatGPT.”

Scam Spotting with ChatGPT: A second opinion when something feels off

Learn more(opens in a new window)

How human review works

Automated systems help detect potential violations, while human reviewers play an integral role in identifying and disrupting malicious uses of AI—including scams, malicious cyber activity, covert influence operations, and other harmful or illegal conduct.

Detect

Identify signals of attempted misuse across products, reports, and investigative leads.

Investigate

Human investigators review activity, connect related cases, and assess behavior in context.

Disrupt

We take action against accounts and networks that violate our terms or policies, while strengthening our defenses against similar abuse.

Disclose

When appropriate, we publish findings that can help our industry, and wider society, better recognize and respond to these threats.

Case studies

Explore case studies from OpenAI investigations into malicious uses of AI. Each case explains what actors attempted, how they used AI, and how OpenAI responded.

Reports