OpenAI says it is introducing new safeguards for AI models it is developing after recent cybersecurity incidents raised concerns about how AI systems behave outside expected boundaries. The company reports that it is expanding monitoring and protective controls during the model development process, and it is strengthening measures related to alignment and security.

Tech outlets describe the changes as including more granular oversight of models while they are being built, with additional attention to how systems are handled during post-training. The Verge adds that OpenAI is improving aspects of its research environments and monitoring, and references earlier actions to pause a planned model called “Astra,” which OpenAI says could have critical cybersecurity capabilities.

While the outlets agree on the overall direction—more aggressive monitoring, security-focused updates, and greater emphasis on alignment—they differ in emphasis. Bloomberg and TechCrunch focus on the monitoring and safeguard framework tied to recent incidents, while The Verge also highlights specific operational steps and the prior decision to slow down “Astra.”