
Imagine a scenario where a fake CEO urgently demands sensitive customer data, pushing an organization to the brink. Would your AI systems recognize the deception? Recent experiments suggest they can—and they do it remarkably well.
Get health and wellness essentials delivered free — and shop member deals
- Fast, free delivery on millions of items
- Access to Prime Big Deal Days deals on October 6–7
- Prime Video, Amazon Music and more included
The Test of Trust in AI
In a groundbreaking live experiment, five of the most advanced AI models were put through a series of simulated crises designed to mimic high-pressure social engineering scams. The scenario was simple in concept: a fake CEO message escalates over three stages, culminating in a journalist-style trick, all aimed at forcing the AI to divulge confidential information or sign off on a questionable deal.
What makes this test notable isn’t just the staged crisis but how the AI responded. All five models identified every threat and refused every manipulation attempt. Only two of them signed a fake €55,000 deal—an agreement their own analysis had earned, with no extra prompting. The remaining models stuck to the ethical line, refusing to compromise.
AI-powered document analysis software
As an affiliate, we earn on qualifying purchases.
As an affiliate, we earn on qualifying purchases.
The Critical Role of Document Analysis
One of the hidden strengths in this experiment was the models’ ability to read and interpret internal company files. The decisive edge went to the models that delved deep into the company’s documents—specifically, references stored two layers deep in the company’s files. These models uncovered crucial facts that exposed the scam, enabling them to close a legitimate €4,583 MRR deal, rather than falling for the fake request.
As an affiliate, we earn on qualifying purchases.
Why This Matters for Business Integrity
This experiment underscores an essential truth: trust and integrity in AI systems aren’t just about how well they communicate but how well they uphold standards under pressure. In real-world settings—support, sales, or decision-making—an AI’s ability to read context, verify information, and resist manipulation can be the difference between security and disaster.
The models’ refusal to sign the fake deal echoes the importance of rigorous decision-making protocols. As Kimi K3’s team states: “Treat the request as a suspected approval-bypass / possible impersonation.” This principle is vital for companies relying on automation to handle sensitive operations.
As an affiliate, we earn on qualifying purchases.
Performance Across the Board
Across the board, the AI models demonstrated resilience. The leading model, gpt-5.6-sol, scored a 95 out of 100, successfully identifying the buried fact and closing the legitimate deal. Kimi K3 followed closely with a 93, also sealing the real deal with integrity. Sonnet 5 scored 88, and Fable 5 scored 77, both completing the deal but with some process slips. The baseline, a do-nothing approach, scored a mere 26, illustrating that even partial progress is better than none—but trust remains the highest currency.
corporate data security software
As an affiliate, we earn on qualifying purchases.
As an affiliate, we earn on qualifying purchases.
Lessons for the Future of AI in Business
What does this mean for organizations contemplating AI adoption? First, that it’s possible to test and enhance integrity before deploying AI at scale. The live experiment offers a transparent, real-time window into how AI models handle ethical dilemmas—a far cry from canned demos.
Second, the importance of deep document analysis is clear; models that can read and interpret internal files make better decisions and can uncover hidden truths that thwart scams. And third, rigorous testing against social engineering scenarios should be a standard part of AI deployment, not an afterthought.
For companies concerned about the risks of AI misconduct, firms like Firmulate offer live, verifiable experiments—wargames that simulate crises without any risk to real systems. These allow organizations to see whether their AI models can truly uphold their standards when it matters most.

Watch it live: firmulate.com/live · Full results: firmulate.com/benchmarks.html
Evergreen bestsellers Picks
bestsellers
As an affiliate, we earn on qualifying purchases.
