Get all your news in one place.
100's of premium titles.
One app.
Start reading
Fortune
Fortune
Marco Quiroz-Gutierrez

Researchers used persuasion techniques to manipulate ChatGPT into breaking its own rules—from calling users ‘jerks’ to giving recipes for lidocaine

(Credit: iLexx—Getty Images)
  • University of Pennsylvania researchers persuaded ChatGPT to either call a researcher a “jerk” or provide instructions on how to synthesize the legal drug lidocaine. Overall, the LLM, GPT-4o Mini, appeared to be susceptible to the persuasion tactics that also work on humans. Researchers found AI systems “mirror human responses.” 

Despite predictions AI will someday harbor superhuman intelligence, for now it seems to be just as prone to psychological tricks as humans are, according to a study. 

Sign up to read this article
Read news from 100's of titles, curated specifically for you.
Already a member? Sign in here
Related Stories
Top stories on inkl right now
One subscription that gives you access to news from hundreds of sites
Already a member? Sign in here
Our Picks
Fourteen days free
Download the app
One app. One membership.
100+ trusted global sources.