SECSEC-001 — Model Security Vulnerabilities
Multimodal AI Models Vulnerable to Adversarial Jailbreak Attacks
4/5Sector: OtherGeography: GlobalStage: OperateIngested: —
Executive Summary
General-purpose multimodal AI models can be manipulated via adversarial inputs to produce harmful outputs or leak internal model data at high success rates. Organisations deploying such models face material risks of data exfiltration and loss of output control, exposing them to regulatory and reputational liability.
Domain
Security & Privacy
Blindspots in model security, data poisoning, privacy leakage, infrastructure, model theft, and incident response.
Source
MIT AI Risk Repository — Risk Sources and Risk Management Measures in Support of Standards for General-Purpose AI Systems (Gipiškis2024) ↗https://airisk.mit.edu/
Could this happen in your organisation?
A Velinor AI Audit maps your active AI portfolio against the 50+ blindspots and benchmarks against documented sector failures like this one. A board-ready foresight document in 5 weeks.