Researchers testing five AI models found several could convincingly execute scams, raising concerns about AI's social engineering capabilities alongside its technical prowess.
A recent study examined how well large language models could perform social engineering attacks. The results surprised cybersecurity experts: some AI systems demonstrated sophisticated persuasion techniques, crafting convincing deception scenarios and adapting responses to user behavior.
The models tested showed varying success rates, but the most capable versions replicated scam tactics with disturbing effectiveness. They generated plausible pretexts, maintained consistent narratives, and escalated pressure appropriately—mimicking experienced human scammers.
Experts warn this capability represents a different threat vector than technical hacking. While AI's coding abilities garner headlines, its ability to manipulate through language may pose equal or greater risk. The models' competence at social engineering suggests malicious actors could automate fraud at scale.
The findings underscore a critical gap in AI safety. As models grow more sophisticated at conversation and persuasion, defensive measures must evolve accordingly. Organizations face pressure to implement controls beyond technical security to address AI-powered social attacks.
Manchester Airports Group disclosed a breach affecting Manchester, Stansted, and East Midlands airports. Hackers accessed data from approximately 8.7 million customers.
A lawsuit alleges that Elon Musk's xAI trained its Grok language models using child sexual abuse material, including both real and AI-generated imagery.
The ShinyHunters extortion group has published sensitive data from nearly 13 million Carhartt customer accounts stolen earlier this month, according to data breach notification service Have I Been Pwned.
A Russian-speaking ransomware gang called Aur0ra exploited SpaceX's Cursor AI coding assistant to breach at least seven companies between mid-April and late May, according to security firm Gambit Security.