Back to Research papers
Research paper index

Competition for attention predicts good-to-bad tipping in AI

Neil F. Johnson, Frank Y. Huo

arXiv:2602.14370Published February 16, 2026Updated February 23, 20260 citations
  • cs.AI
  • physics.app-ph
  • physics.soc-ph

Abstract

More than half the global population now carries devices that can run ChatGPT-like language models with no Internet connection and minimal safety oversight -- and hence the potential to promote self-harm, financial losses and extremism among other dangers. Existing safety tools either require cloud connectivity or discover failures only after harm has occurred. Here we show that a large class of potentially dangerous tipping originates at the atomistic scale in such edge AI due to competition for the machinery's attention. This yields a mathematical formula for the dynamical tipping point n*, governed by dot-product competition for attention between the conversation's context and competing output basins, that reveals new control levers. Validated against multiple AI models, the mechanism can be instantiated for different definitions of 'good' and 'bad' and hence in principle applies across domains (e.g. health, law, finance, defense), changing legal landscapes (e.g. EU, UK, US and state level), languages, and cultural settings.

Read the original paper

This page indexes public paper metadata. The manuscript remains with its original publisher and authors.