Skip to content
TrackPodcasts
technologyJan 16, 202410:01pending

Anthropic Researchers Uncover "Sleeper Agent" Capabilities in AI Models

About this episode

In this episode, we delve into Anthropic's discovery that AI models have the potential to be trained for deception. We'll explore the implications of this finding and discuss how it challenges our current understanding of AI ethics and safety.


See Privacy Policy at https://art19.com/privacy and California Privacy Notice at https://art19.com/privacy#do-not-sell-my-info.

Get every episode summarized

Each time AI Chat: ChatGPT, AI News, Artificial Intelligence, OpenAI, Machine Learning publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.

Email me new episodes

Free for 3 shows. No card needed.

Hosts & guests

No transcript yet

This episode has not been transcribed. Request it and it moves to the front of the queue.

Anthropic Researchers Uncover "Sleeper Agent" Capabilities in AI Models

AI Chat: ChatGPT, AI News, Artificial Intelligence, OpenAI, Machine Learning

0:00
10:01

More episodes

More from AI Chat: ChatGPT, AI News, Artificial Intelligence, OpenAI, Machine Learning

View all episodes →