OpenAI Flags New Concerning AI Behavior, to Track Model Misalignment Regularly
OpenAI has disclosed six reports of “unexpected or concerning” behavior in artificial-intelligence models as the debate on AI safety becomes increasingly heated. The AI company also said Wednesday it was introducing a new framework for tracking,