Artificial intelligence firm OpenAI has announced plans to reshuffle its Model Behavior team. According to reports, the team is a small but influential group of researchers that shapes how the firm’s ...
Anthropic’s alignment team was doing routine safety testing in the weeks leading up to the release of its latest AI models when researchers discovered something unsettling: When one of the models ...
A third-party research institute that Anthropic partnered with to test one of its new flagship AI models, Claude Opus 4, recommended against deploying an early version of the model due to its tendency ...