flash on Nostr: ⚡️🤖 NEW - Researchers have discovered that ChatGPT can be “manipulated” ...
⚡️🤖 NEW - Researchers have discovered that ChatGPT can be “manipulated” using persuasion techniques that are more than 40 years old.
An initial study led by researchers from Wharton and Robert Cialdini himself tested 28,000 conversations with GPT-4o mini, applying the seven principles of persuasion popularized in his 1984 book *Influence*.
When faced with requests that the model was supposed to refuse, these techniques increased its average compliance rate from 33.3% to 72%.
The most striking technique was “commitment”: getting the model to agree to a relatively innocuous initial request before gradually moving on to the problematic request.
In some tests, compliance reached 100%.
Simply changing the way a request was phrased was enough to significantly alter the model’s behavior.
In May 2026, the researchers published a much larger study in PNAS, featuring 126,000 conversations with GPT-5 mini, Claude Haiku 4.5, and Gemini 3 Flash.
Recent models are more resistant, but the effect still exists: when faced with the tested prompts, persuasion increased compliance from 35.3% to 51.3%.
Published at
2026-10-02 18:30:14 UTCEvent JSON
{
"id": "c76d3544b674cb5067a5a164b6d09f1efd343836e25d7aa040c283f07193ad4b",
"pubkey": "4d7842051782e0d3feb034d150adc2b6bae4ee3b49786793bffa468b6f5b96b3",
"created_at": 1790965814,
"kind": 1,
"tags": [
[
"client",
"Primal iOS"
]
],
"content": "⚡️🤖 NEW - Researchers have discovered that ChatGPT can be “manipulated” using persuasion techniques that are more than 40 years old.\n\nAn initial study led by researchers from Wharton and Robert Cialdini himself tested 28,000 conversations with GPT-4o mini, applying the seven principles of persuasion popularized in his 1984 book *Influence*.\n\nWhen faced with requests that the model was supposed to refuse, these techniques increased its average compliance rate from 33.3% to 72%.\n\nThe most striking technique was “commitment”: getting the model to agree to a relatively innocuous initial request before gradually moving on to the problematic request.\nIn some tests, compliance reached 100%.\n\nSimply changing the way a request was phrased was enough to significantly alter the model’s behavior. \n\nIn May 2026, the researchers published a much larger study in PNAS, featuring 126,000 conversations with GPT-5 mini, Claude Haiku 4.5, and Gemini 3 Flash.\n\nRecent models are more resistant, but the effect still exists: when faced with the tested prompts, persuasion increased compliance from 35.3% to 51.3%. \nhttps://blossom.primal.net/1ac882832eb8b27a0851590f97b1f2da8b13e80bf3ea9b940ecc2ace023ea2ee.jpg\nhttps://blossom.primal.net/3dafba42156e0c67ea3ab30c0414f7cab5776f178384287b53e3cd33dd562164.jpg",
"sig": "3aab1b49a2d55adcab54ec895fc96c2d5afbf54850b7fc57f492d283fcff445157cb4e2d0b178404b8e4043ed6dc4fc494c78e081b23f10dd76a2888ab0df030"
}