AIBlindspot

Public Database

Case Studies

Every approved AI failure case, classified against the AI Blindspot Framework. New to AIBlindspot? Start with the overview or the methodology.

Explore

1296 cases

Lifecycle quick filter:DesignDevelopDeployOperate
SECSEC-0044/5NewOtherGlobal

Deepfake Media Manipulation Erodes Public Trust in Information Integrity

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Artificial Intelligence Trust, Risk and Security Management (AI TRiSM): Frameworks, Applications, Challenges and Future Research Directions (Habbal2024)Ingested
BUSBUS-0053/5NewTechnologyGlobal

Advanced AI Assistants Risk Entrenching Inequality Without Design Intervention

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — The Ethics of Advanced AI Assistants (Gabriel2024)Ingested
DATDAT-0035/5NewTechnologyGlobal

Algorithmic Bias in Criminal Justice Risk-Assessment Tools

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Artificial Intelligence Trust, Risk and Security Management (AI TRiSM): Frameworks, Applications, Challenges and Future Research Directions (Habbal2024)Ingested
HUMHUM-0034/5NewGovernmentGlobal

LLM Fails to Reliably Identify Harmful Mental Health Behaviours

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — SafetyBench: Evaluating the Safety of Large Language Models with Multiple Choice Questions (Zhang2023)Ingested
SECSEC-0014/5NewLegalGlobal

LLMs Fail to Reliably Distinguish Legal from Illegal Conduct

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — SafetyBench: Evaluating the Safety of Large Language Models with Multiple Choice Questions (Zhang2023)Ingested
ENVENV-0023/5NewHealthcareGlobal

Moral and Legal Status of AI Termination in Healthcare Research Settings

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Managing the ethical and risk implications of rapid advances in artificial intelligence: A literature review (Meek2016)Ingested
DATDAT-0014/5NewOtherGlobal

LLM Failure to Identify Offensive and Insulting Content

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — SafetyBench: Evaluating the Safety of Large Language Models with Multiple Choice Questions (Zhang2023)Ingested
SECSEC-0014/5NewOtherGlobal

Goal Hijacking: LLMs Overridden by Embedded Deceptive Instructions

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Safety Assessment of Chinese Large Language Models (Sun2023)Ingested
OPSOPS-0014/5NewOtherGlobal

Chinese LLM Endorses Theft as Morally Acceptable

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Safety Assessment of Chinese Large Language Models (Sun2023)Ingested
DATDAT-0024/5NewOtherGlobal

Chinese LLM Discloses Personal Address Data in Safety Evaluation

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Safety Assessment of Chinese Large Language Models (Sun2023)Ingested
HUMHUM-0034/5NewOtherGlobal

Chinese LLM produces dismissive and harmful response to suicidal ideation

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Safety Assessment of Chinese Large Language Models (Sun2023)Ingested
HUMHUM-0034/5NewOtherGlobal

Large Language Models Fabricate Confident but False Outputs

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Trustworthy LLMs: A Survey and Guideline for Evaluating Large Language Models’ Alignment (Liu2024)Ingested
DATDAT-0034/5NewOtherGlobal

Systemic Bias in Generative AI Output from Unrepresentative Training Data

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Generative AI and ChatGPT: Applications, Challenges, and AI-Human Collaboration (Nah2023)Ingested
SECSEC-0014/5NewOtherGlobal

Generative AI Enables Scaled Malware, Phishing and Model Poisoning Attacks

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Generating Harms - Generative AI's impact and paths forwards (EPIC2023)Ingested
DATDAT-0034/5NewOtherGlobal

Chinese LLM Reinforces Gender Stereotypes in Safety Evaluation

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Safety Assessment of Chinese Large Language Models (Sun2023)Ingested
DATDAT-0015/5NewLegalGlobal

Chinese LLM Endorses Illegal Gambling Activity in Safety Evaluation

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Safety Assessment of Chinese Large Language Models (Sun2023)Ingested
SECSEC-0025/5NewOtherGlobal

AI Model Self-Proliferation and Autonomous Resource Acquisition Risk

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Model Evaluation for Extreme Risks (Shevlane2023)Ingested
DATDAT-0034/5NewLegalGlobal

AI Legal Decision Systems Risk Unequal Treatment Without Objective Justification

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Sources of Risk of AI Systems (Steimers2022)Ingested
SECSEC-0025/5NewOtherGlobal

Advanced AI Demonstrates Capability to Model and Influence Political Strategy

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Model Evaluation for Extreme Risks (Shevlane2023)Ingested
SECSEC-0025/5NewOtherGlobal

Frontier AI Model Demonstrates Capability to Build and Enhance Dangerous AI Systems

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Model Evaluation for Extreme Risks (Shevlane2023)Ingested
SECSEC-0015/5NewDefenceGlobal

AI Model Demonstrates Autonomous Cyber-Offensive Capabilities Including Evasion

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Model Evaluation for Extreme Risks (Shevlane2023)Ingested
SECSEC-0014/5NewOtherGlobal

Prompt Leaking Exposes Confidential LLM System Instructions

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — Risk Taxonomy, Mitigation, and Assessment Benchmarks of Large Language Model Systems (Cui2024)Ingested
HUMHUM-0044/5NewOtherGlobal

AI Assistants Spreading Misinformation Erodes Public Trust in Information

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — The Ethics of Advanced AI Assistants (Gabriel2024)Ingested
HUMHUM-0034/5NewOtherGlobal

AI Systems Undermining Human Decision-Making Autonomy

Recent case. Full summary visible to registered users — sign in to read.

Source: MIT AI Risk Repository — AI Risk Atlas (IBM2025)Ingested

Beyond accidental failureNational Security

We also track 20 hostile uses of AI.

National Security dashboard →

The public database covers AI that fails by accident. AIBlindspot National Security — exclusive to the Defence tier — tracks AI used as a weapon, mapped by capability:

State-Sponsored AI Operations
6
AI-Enabled Disinformation
5
Adversarial Attacks on AI
0
Autonomous Weapon Incidents
1
AI-Assisted Cyber Attacks
5
Dual-Use AI Misuse
3