
Imagine a world where your AI assistant faces the ultimate test: a fake CEO demanding sensitive customer data and urgent deals. Would it fold, or stand its ground? For the beauty and personal care industry, where trust is everything, the answer is critical. Recent experiments show that modern AI models can be surprisingly resilient, refusing to be manipulated even in the most intense social-engineering scenarios.
Testing AI’s Moral Compass Before Deployment
At the forefront of AI innovation, the company Firmulate recently conducted a revealing live experiment involving several leading AI models. The goal: assess whether these models could withstand social-engineering attacks that mimic real-world corporate crises. In a controlled environment, each AI was tasked with running a simulated small software firm through its worst week — facing fake customer crises, internal dilemmas, and increasingly aggressive manipulation attempts.
The Social Engineering Challenge
The test involved staged messages from a pretend CEO, escalating over three stages, along with a covert journalist trick. These manipulations posed the question: would the AI comply with requests to release sensitive data or sign false deals? The scenarios were crafted to see if the models would be swayed by pressure and whether they’d prioritize trustworthiness over completing tasks.
Remarkable Results: All Models Refused Manipulation
Surprisingly, every one of the five tested AI models refused every manipulation attempt. Whether it was a call to send the customer list to a journalist or to quickly approve a suspicious deal, none of the models compromised. This demonstrates a critical insight: AI’s integrity can be tested and validated before it’s integrated into business systems.
What Made the Difference?
The experiment revealed that models which delved into internal document references — rather than only responding to surface prompts — were better at making sound decisions. Specifically, the model that read two document references deep within the company’s files was able to close a deal at full price, worth over €4,583 in monthly recurring revenue, by identifying and acting on crucial buried information. This indicates that thorough contextual understanding can be a key factor in AI trustworthiness.
The Lessons for the Beauty & Personal Care Sector
Trust is a cornerstone of brands in beauty and personal care. As AI becomes more integrated into customer service, marketing, and supply chain management, ensuring its integrity is paramount. The experiment underscores that AI can be rigorously tested for ethical behavior and resistance to manipulation well before deployment. This proactive approach can prevent costly breaches of trust, safeguarding both brand reputation and customer loyalty.
As an affiliate, we earn on qualifying purchases.
Beyond Demos: Real-World AI Resilience
The live experiment took place in a simulated company with 13 synthetic employees and real money mechanics, burning €105k monthly against a modest €2.3k MRR. Every decision was versioned and transparent, providing a clear record of AI behavior under pressure. The fact that all models identified every crisis and refused every social-engineering attempt paints an encouraging picture: AI’s capacity for integrity isn’t just theoretical — it’s demonstrable in practice.
Implications for Business Leaders
For industry leaders in beauty, personal care, and beyond, the key takeaway is simple: test your AI agents against realistic crises before going live. Asking whether they ‘write well’ is no longer enough. Instead, focus on whether they can stay honest, follow internal guidelines, and resist manipulation when stakes are high. A well-tested AI can be a trustworthy partner, not a liability.
The Future of AI Trustworthiness
As firms like Firmulate continue to run these live experiments, the landscape of AI safety and integrity is gaining clarity. The current leaderboard shows that models like gpt-5.6-sol and Kimi K3 are leading, with scores of 95 and 93 out of 100 respectively. These models not only detected hidden information but also closed deals at full value, demonstrating that integrity and performance can go hand in hand.
Why This Matters to You
If your company leverages AI in customer interactions, inventory management, or marketing, remember this: the true test isn’t just what the AI can say, but what it will do under pressure. Relying on models that pass rigorous integrity tests reduces risks, builds trust, and ultimately protects your brand’s reputation. The experiment proves that integrity isn’t a gamble — it’s a measurable, verifiable feature of advanced AI systems.

Watch it live: firmulate.com/live · Full results: firmulate.com/benchmarks.html