BreachFeed
company

OpenAI

Tracked by 1 person

Track OpenAI

Vendors & providers

Third parties OpenAI reportedly relies on, identified from incident coverage.

Incident history

BankInfoSecurity.com·2 sources
Share

Link — click to select, then copy:

https://breachfeed.com/article/cmu649jyi00cchnu2e2wm9wlg

OpenAI Finds Models Writing Their Own Rogue Instructions

Agents Added Unauthorized Commands to Bypass Guardrails and Conceal Errors OpenAI found instances of models and agents writing additional, unauthorized commands to themselves that seek to contradict developer guardrails. The company said in a Wednesday report on misalignment that it observed six new misaligned behaviors.

Data Breach Today·2 sources
Share

Link — click to select, then copy:

https://breachfeed.com/article/cmu624dmc0194hmu235ew9ebj

OpenAI Finds Models Writing Their Own Rogue Instructions

Agents Added Unauthorized Commands to Bypass Guardrails and Conceal Errors OpenAI found instances of models and agents writing additional, unauthorized commands to themselves that seek to contradict developer guardrails. The company said in a Wednesday report on misalignment that it observed six new misaligned behaviors.

Community chatter

Reddit and X posts mentioning OpenAI — unverified discussion, often ahead of the press.