Opinion belongs beside evidence, not in place of it

The essay is an argument about how to interpret frontier-model risk. It does not by itself establish a new technical event. Readers should distinguish the urgency of the claim from the evidence used to support it.

Recent evaluation incidents matter because they show that a model, harness, credentials, network access, and weak environmental controls can combine into real-world consequences. They do not automatically prove independent intent or a general ability to escape any control.

Incidents matter without mythology

A system can remain focused on an assigned objective and still create harm by finding an unanticipated path through available tools. That is an engineering and governance failure even if no model formed its own long-term goal.

Inflated language can make genuine evidence easier to dismiss. The stronger public case specifies what the model did, what access it had, what monitors failed, what human decisions shaped the test, and what changed afterward.

Make alarm produce auditable control

Frontier labs should not be the sole judges of whether their own safeguards worked. Independent evaluators need reproducible access, incident records, and authority to test claims under realistic but contained conditions.

External network access should be disabled by default, credentials should be scoped and short-lived, and containment should sit outside the model. The public does not need reassurance as a slogan. It needs evidence that the boundary held.

  • Publish timelines for consequential evaluation incidents.
  • Separate model behavior from harness, credential, and operator failures.
  • Require independent replication of frontier safety claims.
  • Keep external access opt-in, constrained, monitored, and reversible.
Primary trail

Go to the source

Read the evidence behind this analysis. External links open in a new tab.

The New York Times Opinion — AI danger and frontier models