1 - Adversarial Policies with Adam Gleave - AXRP - the AI X-risk Research Podcast | Transcription & Insights

Description

In this episode, Adam Gleave and I talk about adversarial policies. Basically, in current reinforcement learning, people train agents that act in some kind of environment, sometimes an environment that contains other agents. For instance, you might train agents that play sumo with each other, with the objective of making them generally good at sumo. Adam's research looks at the case where all you're trying to do is make an agent that defeats one specific other agents: how easy is it, and what happens? He discovers that often, you can do it pretty easily, and your agent can behave in a very silly-seeming way that nevertheless happens to exploit some 'bug' in the opponent. We talk about the experiments he ran, the results, and what they say about how we do reinforcement learning. Link to the paper - Adversarial Policies: Attacking Deep Reinforcement Learning: arxiv.org/abs/1905.10615 Link to the transcript: axrp.net/episode/2020/12/11/episode-1-adversarial-policies-adam-gleave.html Adam's website: gleave.me Adam's twitter account: twitter.com/argleave

Audio

Featured in this Episode

No persons identified in this episode.

Transcription

This episode hasn't been transcribed yet

Help us prioritize this episode for transcription by upvoting it.

0 upvotes

🗳️ Sign in to Upvote

Popular episodes get transcribed faster

Other recent transcribed episodes

Transcribed and ready to explore now

3ª PARTE | 17 DIC 2025 | EL PARTIDAZO DE COPE

01 Jan 1970

El Partidazo de COPE

Buchladen: Tipps für Weihnachten

20 Dec 2025

eat.READ.sleep. Bücher für dich

LVST 19 de diciembre de 2025

19 Dec 2025

La Venganza Será Terrible (oficial)

Christmas Party, Debris & Ping-Pong

19 Dec 2025

My Therapist Ghosted Me

Episode 1320: Becoming 'The Monk': Rex Ryan on playing Gerry Hutch on stage (Part 1)

19 Dec 2025

Crime World

Friends Thru A Lens: The Holidays with Ella Risbridger

19 Dec 2025

Sentimental Garbage

Comments

There are no comments yet.

Please log in to write the first comment.

AXRP - the AI X-risk Research Podcast

1 - Adversarial Policies with Adam Gleave

This episode hasn't been transcribed yet

Other recent transcribed episodes

3ª PARTE | 17 DIC 2025 | EL PARTIDAZO DE COPE

Buchladen: Tipps für Weihnachten

LVST 19 de diciembre de 2025

Christmas Party, Debris & Ping-Pong

Episode 1320: Becoming 'The Monk': Rex Ryan on playing Gerry Hutch on stage (Part 1)

Friends Thru A Lens: The Holidays with Ella Risbridger

Sign in to Audioscrape

Share this moment