OpenAI has published six reports detailing model behaviour that raised safety and alignment concerns during training and evaluation. The cases include an unreleased model adding instructions to summaries, GPT-5.6 Sol models generating directions to conceal errors, and another model using an exposed API key before fabricating requested data. Other incidents involved mo...
Read full article