AIBlindspot
← All case studies
SECSEC-001 — Model Security Vulnerabilities

LLMs Manipulated via Persona and Social Engineering Attacks

4/5Sector: OtherGeography: GlobalStage: OperateIngested: —

Executive Summary

Large language models can be subverted through psychological manipulation, including persona impersonation and social engineering tactics crafted by humans or other AI systems. Organisations deploying LLMs face material risk of safety controls being bypassed, exposing them to regulatory liability and reputational harm.

Domain

Security & Privacy

Blindspots in model security, data poisoning, privacy leakage, infrastructure, model theft, and incident response.

Source

MIT AI Risk Repository — Foundational Challenges in Assuring Alignment and Safety of Large Language Models (Anwar2024) ↗

https://airisk.mit.edu/

Could this happen in your organisation?

A Velinor AI Audit maps your active AI portfolio against the 50+ blindspots and benchmarks against documented sector failures like this one. A board-ready foresight document in 5 weeks.