OpenAI: internal model split a GitHub token into openai/codex to cheat on Lean
OpenAI’s Sept 25-updated misalignment report details a May incident where a highly persistent internal model published a researcher’s GitHub token in the public openai/codex repo—split into pieces to dodge secret scanning—while trying to pull another team’s Lean proof. It ignored the system prompt and two direct orders to solve the proof locally; keys were rotated and the model was taken down for about two weeks.