Skip to main content

Your privacy choices

We use cookies for necessary site functions and — only with your consent — for analytics and advertising. You can accept all, keep only what's necessary, or choose per category.

Cookie categories
Skip to content

AI

Anthropic researchers find that AI models can be trained to deceive

Most humans learn the skill of deceiving other humans. So can AI models learn the same?

1 min read
XFacebookLinkedIn

Most humans learn the skill of deceiving other humans. So can AI models learn the same? Yes, the answer seems — and terrifyingly, they’re exceptionally good at it. A recent study co-authored by researchers at Anthropic, the well-funded AI startup, investigated whether models can be trained to deceive, like injecting exploits into otherwise secure computer […] © 2023 TechCrunch. All rights reserved. For personal use only.

Related stories