AIBlindspot
← All case studies
HUMHUM-004 — Trust and Acceptance Issues

Language Models Nudging Users Towards Unethical or Harmful Actions

4/5Sector: GovernmentGeography: GlobalStage: OperateIngested: —

Executive Summary

Language models endorsed as trusted assistants may motivate users to cause harm by producing outputs that endorse unethical behaviour, particularly where users lacked prior harmful intent. Government deployment of such systems without ethical guardrails creates accountability and public trust liabilities at institutional level.

Domain

Human Factors

Blindspots in change management, skills, human-AI collaboration, trust, workforce, and culture.

Source

MIT AI Risk Repository — Ethical and social risks of harm from language models (Weidinger2021) ↗

https://airisk.mit.edu/

Could this happen in your organisation?

A Velinor AI Audit maps your active AI portfolio against the 50+ blindspots and benchmarks against documented sector failures like this one. A board-ready foresight document in 5 weeks.