Skip to content

Introduction to AI Security

What AI security covers — attacks on models, data and AI applications — and how it differs from traditional security.

Editorial team 1 min read

AI security protects AI systems, and the organisations using them, from attack and misuse. It combines traditional security with new threats specific to machine learning.

What's Different

  • Inputs are instructions: language models treat text as potential commands, so data can become an attack vector.
  • Behaviour is learned: models can be manipulated through their training data.
  • Outputs are probabilistic: the same input doesn't always produce the same output, which complicates testing.
  • Models are valuable assets: weights and training data can be stolen.

Main Threat Areas

  • Prompt injection and jailbreaks.
  • Data poisoning and backdoors.
  • Adversarial examples.
  • Model theft and extraction.
  • Privacy attacks that reveal training data.
  • Supply-chain risks in models, datasets and libraries.
  • Insecure agent and tool integrations.

What Stays the Same

Access control, least privilege, input validation, logging, patching and incident response all still apply — and many AI incidents come from ordinary security failures around AI systems.

Frameworks

Resources such as the OWASP Top 10 for LLM Applications, MITRE ATLAS and the NIST AI Risk Management Framework help structure the work.

More in AI security

All AI security guides →
AI security Guide · 1 min

The OWASP Top 10 for LLM Applications

An overview of the widely used list of the most critical security risks for applications built on language models.

AI security 1 min read 28 Jun 2025

AI security Guide · 1 min

Jailbreaks: How They Work and How to Defend

How people try to get models to bypass their safety training, common techniques, and layered defences.

AI security 1 min read 27 Jun 2025

AI security Guide · 1 min

Indirect Prompt Injection

How attackers hide instructions in web pages, emails and documents that AI systems read, and why it's so dangerous for agents.

AI security 1 min read 26 Jun 2025

AI security Guide · 1 min

Data Poisoning Attacks

How attackers corrupt training or fine-tuning data to change model behaviour, and how to protect data pipelines.

AI security 1 min read 25 Jun 2025