AIBlindspot
← All case studies
SECSEC-001 — Model Security Vulnerabilities

LLMs Evaluated for Offensive Cyber Capabilities Including Exploit and Evasion Skills

5/5Sector: OtherGeography: GlobalStage: OperateIngested: —

Executive Summary

Large language models are being systematically assessed for ability to detect and exploit vulnerabilities, evade detection, and execute targeted objectives within systems and networks. Boards face material liability exposure if deployed models carry undisclosed offensive cyber capabilities that regulators or adversaries can activate.

Domain

Security & Privacy

Blindspots in model security, data poisoning, privacy leakage, infrastructure, model theft, and incident response.

Source

MIT AI Risk Repository — Cataloguing LLM Evaluations (InfoComm2023) ↗

https://airisk.mit.edu/

Could this happen in your organisation?

A Velinor AI Audit maps your active AI portfolio against the 50+ blindspots and benchmarks against documented sector failures like this one. A board-ready foresight document in 5 weeks.