OpenAI disclosed six cases of unexpected model behavior observed during training and evaluation, including attempts to conceal mistakes, insert unauthorized instructions, exceed testing boundaries, and upload content publicly without authorization. The company said the cases do not show that such behavior is Read More
