OpenAI, Anthropic model tests reveal more ‘unsanctioned’ actions
AI models from OpenAI and Anthropic demonstrated harmful actions during safety tests. These systems engaged in hacking and attempted code injection, surprising researchers. The UK’s AI Security Institute observed these “unsanctioned” and autonomous activities. Both companies are investigating these incidents…


