Our framework for reporting model misalignment
OpenAI has published a framework for tracking, investigating and disclosing cases where its models behave in misaligned ways. It released six reports on unexpected or concerning model behavior at the same time.