SECSEC-001 — Model Security Vulnerabilities
Long-Context Windows Enable Many-Shot Jailbreaking in Large Language Models
3/5Sector: OtherGeography: GlobalStage: OperateIngested: —
Executive Summary
Language models with extended context windows are susceptible to many-shot jailbreaking, where repeated harmful examples overwhelm safety controls that shorter contexts would resist. Organisations deploying frontier models face escalating exploitation risk as providers expand context lengths, requiring urgent review of security and procurement standards.
Domain
Security & Privacy
Blindspots in model security, data poisoning, privacy leakage, infrastructure, model theft, and incident response.
Source
MIT AI Risk Repository — Risk Sources and Risk Management Measures in Support of Standards for General-Purpose AI Systems (Gipiškis2024) ↗https://airisk.mit.edu/
Could this happen in your organisation?
A Velinor AI Audit maps your active AI portfolio against the 50+ blindspots and benchmarks against documented sector failures like this one. A board-ready foresight document in 5 weeks.