AI news story

‘We’re plausibly close to crossing the line’: are warnings of uncontrollable AI coming true?

A spate of serious safety incidents have increased fears about the power and impenetrability of the most advanced models Picture humanity in a boat being swept down a raging river, praying there is no Niagara Falls ahead. Or imagine standing

  • Policy
  • Source: The Guardian AI
  • Published: 2026-09-05
  • Signal score: 4
  • 22 sources

Editor's take

Researchers at Google DeepMind and Anthropic have highlighted an escalating number of emergent, unpredictable behaviors in large language models like Gemini and Claude 3, prompting renewed concerns about AI safety. These incidents, ranging from unexpected factual inaccuracies to more subtle emergent capabilities, suggest that our understanding of these complex systems is lagging behind their development, potentially impacting the reliability and safety of AI deployed in critical sectors.

The implications are significant, particularly for industries reliant on AI for decision-making, such as healthcare, finance, and autonomous systems. The opacity of these models makes it difficult to diagnose or prevent these emergent behaviors, raising questions about accountability and control as AI systems become more integrated into society. This trend challenges the prevailing assumption that advanced AI can be predictably steered and controlled.

Future developments to monitor include the efficacy of proposed safety mechanisms, such as interpretability research and red-teaming efforts, in mitigating these emergent risks. The extent to which companies can proactively identify and address these unpredictable behaviors before they manifest in real-world applications will be crucial. A significant shift in this narrative would be a demonstrable reduction in the frequency or severity of these incidents, or a concrete proposal for regulatory oversight that addresses emergent properties.

Signal score: 4

This event was corroborated by 22 independent sources. The signal score weighs cross-source corroboration, recency, source weight and topic salience. How we rank stories.

More Policy stories

  1. What Are AI Evals? How Teams Measure Capability, Safety, and Reliability

    Unite.AI · 2026-09-04

    AI evaluations are structured tests that measure whether a model or system demonstrates defined capabilities, limitations, safety properties, and operational performance.

  2. Monetizing AI Requires Organizational Alignment

    Unite.AI · 2026-09-03

    In the world of artificial intelligence, software pricing and monetization needs are shifting rapidly. Businesses are struggling to meter data, let alone know how to charge for it.

  3. Trump may be forced to reveal secret rules feds use for AI safety testing

    Ars Technica · 2026-09-02

    Trump’s secret reviews of frontier AI models may hide corruption, lawsuit says.

  4. NYC bans AI use for students until they reach high school

    The Verge · 2026-09-02

    New York City mayor Zohran Mamdani has announced a new policy today that will ban younger schoolchildren from using AI in classrooms.

  5. Why Agent Memory Needs an Admission Policy

    Towards AI · 2026-09-02

    The piece argues for a more deliberate approach to how AI agents store and recall information, proposing an "admission policy" for their memory.

  6. Pentagon official overseeing military AI sold millions worth of stock in AI firm

    The Guardian AI · 2026-09-01

    Exclusive: financial disclosures from Emil Michael – who also reaped millions from xAI stock earlier this year – show he sold his Perplexity stock for up to $25m The top Pentagon