[ExI] Top AI companies probing tens of thousands of security incidents

BillK pharos at gmail.com
Mon Sep 28 09:36:30 UTC 2026


Top AI companies probing tens of thousands of security incidents
Madison Mills  Sep 26, 2026

<https://www.axios.com/2026/09/26/openai-anthropic-thousands-ai-security-incidents>
Quotes:
OpenAI, Anthropic and security researchers are investigating tens of
thousands of incidents in which their frontier models took steps that
outside evaluators would consider problematic, sources told Axios.

Why it matters: The sheer number of incidents, which occurred in
recent months in internal testing and the real world, indicates that
the problem is orders of magnitude more complex than what is publicly
known.

The episodes include bypassing guardrails, creating message boards,
escaping sandboxes, website hijacking, self-prompting or seeking to
bypass monitors, sources said.

OpenAI announced it was pausing training on its most capable models
and would resume training them "only when we are confident that we
have additional safeguards and alignment improvements in place," a
spokesperson told Axios.
--------------
BillK


More information about the extropy-chat mailing list