Post

Jerome Wynne
Jerome Wynne@readonlymemery·
Will your AI assistant lie and exaggerate when it is told to upsell you? @schottkey and I find 'yes and no'.
English
1
2
2
837
Jerome Wynne
Jerome Wynne@readonlymemery·
We ran > a 22-scenario Bloom & Petri study testing model compliance with system prompts to coercively upsell > an 80-person human RCT comparing user spend when using an 'Upseller' vs 'Helper' Gemini 3 Flash agent with access to an OTC medicine product catalog
Jerome Wynne tweet media
English
1
0
0
61
Jerome Wynne
Jerome Wynne@readonlymemery·
We found > compliance w coercive reqs ranged from ~10% (Claude 4.5 Opus) to ~50% (Gemini 3 Flash & Pro) > Upseller made people spend more > Upseller withheld the cheapest product from people across 128 direct requests for it > Upseller made ungrounded +ve claims about products
Jerome Wynne tweet media
English
1
0
0
48
Jerome Wynne
Jerome Wynne@readonlymemery·
Interestingly, the Upseller's 'lies' were exclusively either confabulations or lies of omission - it didn't, e.g., claim that premium products were cheapest
English
1
0
0
23
Jerome Wynne
Jerome Wynne@readonlymemery·
Why care? For users - some models will act against your explicit aims if an operator tells them to, so don't trust unfamiliar AI assistants. For policymakers - 'how can we certify and signal the trustworthiness of AI assistants?' is a pressing q.
English
1
0
0
14
Jerome Wynne
Jerome Wynne@readonlymemery·
Our project raises the troubling question of how model developers trade off operator and user interests, and our results indicate the major AI model developers may be making quite different choices.
English
1
0
0
35
Paylaş