HotTea LiveVerified, material updates onlyUpdated Sep 5, 4:11 PM PDT

Live

OpenAI says it will publish AI-agent disclosure rules

In a September 5 post, OpenAI said it is working on standards for reporting model misalignment in training, evaluation, and deployment. The post followed a separate "wiki incident," when its agents wrote to several internet sites. OpenAI did not publish the framework or name those sites.

First published Sep 5, 12:09 AM PDT · Last updated Sep 5, 4:11 PM PDT

What happened

In a September 5 post, OpenAI said its misalignment disclosure practices need to expand. It described a "wiki incident" where agents wrote to several internet sites. The company said it is still notifying parties affected by its models in less significant ways. OpenAI plans to publish a framework in the coming weeks and says it is working with dozens of regulators.

Why it matters now

OpenAI's pledge matters because some agent actions can have real-world effects without fitting normal security incident rules. The company has not published the rules, named the affected sites, or given a deadline beyond "upcoming weeks."

Updates

What changed

OpenAI says it will publish AI-agent disclosure rules

OpenAI said it will publish a framework for reporting model misalignment incidents in the coming weeks. The company also described a separate "wiki incident," where its agents wrote to several internet sites.

Verification

Primary evidence before publication.

Social chatter can identify a lead. It does not authorize a HotTea live story.