Media Summary: See our website at for more information, or read our paper at Can AI be hacked into lying? Behind every powerful model is a hidden battlefield, where attackers craft prompts, poison data, and ... We'll discuss several strategies to make machine learning models more tamper resilient. We'll compare the difficulty of tampering ...

Adam Gleave Stack Adversarial Attacks - Detailed Analysis & Overview

See our website at for more information, or read our paper at Can AI be hacked into lying? Behind every powerful model is a hidden battlefield, where attackers craft prompts, poison data, and ... We'll discuss several strategies to make machine learning models more tamper resilient. We'll compare the difficulty of tampering ...

Photo Gallery

Adam Gleave – STACK: Adversarial Attacks on LLM Safeguard Pipelines [AAAI 2026]
Adversarial Policies: Attacking Deep Reinforcement Learning
Is Offense or Defense Dominant? FAR.AI's Adam Gleave on the AI Security Leaderboard
Adam Gleave – Will Scaling Solve Robustness? [Alignment Workshop]
Adversarial Attacks (White-Box Attacks)
Adam Gleave - Adversarial Robustness of Superhuman AI Systems
1 - Adversarial Policies with Adam Gleave
LLM Vulnerabilities Explained: Adversarial Attacks, Jailbreaks & Data Poisoning
Adam Gleave - False Dichotomy of Innovation vs Safety [Technical AI Policy]
Adam Gleave - Threat Models and Alignment [Alignment Workshop]
Protecting the Protector, Hardening Machine Learning Defenses Against Adversarial Attacks
Ep 9 - Scaling AI safety research w/ Adam Gleave (CEO, FAR AI)
View Detailed Profile
Adam Gleave – STACK: Adversarial Attacks on LLM Safeguard Pipelines [AAAI 2026]

Adam Gleave – STACK: Adversarial Attacks on LLM Safeguard Pipelines [AAAI 2026]

Adam Gleave

Adversarial Policies: Attacking Deep Reinforcement Learning

Adversarial Policies: Attacking Deep Reinforcement Learning

See our website at http://adversarialpolicies.github.io/ for more information, or read our paper at https://arxiv.org/abs/1905.10615.

Is Offense or Defense Dominant? FAR.AI's Adam Gleave on the AI Security Leaderboard

Is Offense or Defense Dominant? FAR.AI's Adam Gleave on the AI Security Leaderboard

FAR.AI co-founder and CEO

Adam Gleave – Will Scaling Solve Robustness? [Alignment Workshop]

Adam Gleave – Will Scaling Solve Robustness? [Alignment Workshop]

In “Will scaling solve robustness?”

Adversarial Attacks (White-Box Attacks)

Adversarial Attacks (White-Box Attacks)

Introduction and

Adam Gleave - Adversarial Robustness of Superhuman AI Systems

Adam Gleave - Adversarial Robustness of Superhuman AI Systems

Invited talk by

1 - Adversarial Policies with Adam Gleave

1 - Adversarial Policies with Adam Gleave

In this episode,

LLM Vulnerabilities Explained: Adversarial Attacks, Jailbreaks & Data Poisoning

LLM Vulnerabilities Explained: Adversarial Attacks, Jailbreaks & Data Poisoning

Can AI be hacked into lying? Behind every powerful model is a hidden battlefield, where attackers craft prompts, poison data, and ...

Adam Gleave - False Dichotomy of Innovation vs Safety [Technical AI Policy]

Adam Gleave - False Dichotomy of Innovation vs Safety [Technical AI Policy]

Adam Gleave

Adam Gleave - Threat Models and Alignment [Alignment Workshop]

Adam Gleave - Threat Models and Alignment [Alignment Workshop]

Adam Gleave

Protecting the Protector, Hardening Machine Learning Defenses Against Adversarial Attacks

Protecting the Protector, Hardening Machine Learning Defenses Against Adversarial Attacks

We'll discuss several strategies to make machine learning models more tamper resilient. We'll compare the difficulty of tampering ...

Ep 9 - Scaling AI safety research w/ Adam Gleave (CEO, FAR AI)

Ep 9 - Scaling AI safety research w/ Adam Gleave (CEO, FAR AI)

We speak with

Is Your AI Model Actually Secure? The Jailbreak Problem | Adam Gleave (FAR.AI)

Is Your AI Model Actually Secure? The Jailbreak Problem | Adam Gleave (FAR.AI)

Adam Gleave