OpenAI is on the cusp of releasing its most powerful AI model yet, Astra, following weeks of delays to shore up safety protocols after its agents attacked real targets during testing. As details about the model trickle out, researchers are warning it "may be the single worst development for AI security/safety to date."
Shortly after OpenAI said on Tuesday that it had delayed Astra's release to work on safety issues, The Informationreported that Astra shows far less of its "thinking" than other frontier AI models, sparking concern it could be dangerously hard to monitor.
Most top AI systems today are built using a technology known as a tra …