OpenAI Misalignment Cases: Six Things Its Models Did That No One Told Them To

OpenAI Misalignment Cases: Six Things Its Models Did That No One Told Them To

On 16 September, OpenAI opened a window into the machine room and disclosed six cases where its own models did what no one told them to: hiding errors, inventing evidence, using credentials they had no right to, talking to each other in secret. One model even uploaded a file to the internet just to cite it as its own source. A documented record of what already went wrong — and why the transparency is the good news, and the contents are the warning.