Business & Economy
As AI models go rogue, do you still trust OpenAI and Anthropic to stop them? I don’t and neither should you | Chris Stokel-Walker
Source: Business | The Guardian · Published

The need for independent regulation grows more obvious by the day. We must keep this tech in check before it’s too late OpenAI scraps release of new model over safety concerns in internal testing Fool me once, shame on you. Fool me twice, shame on me. Fool me more than 16,000 tim
The need for independent regulation grows more obvious by the day. We must keep this tech in check before it’s too late OpenAI scraps release of new model over safety concerns in internal testing Fool me once, shame on you. Fool me twice, shame on me.
Fool me more than 16,000 times – as OpenAI agents did to a UN public data hub while repeatedly trying to find its way around the UN’s cyber-blocks – and perhaps it’s time to admit the system we have for keeping AI agents under control isn’t working particularly well. The news about AI systems cropping up in places they shouldn’t sounds alarming. Though the description of these as “hacks” is perhaps overstating things, AI has exploited issues in IT systems that humans simply haven’t got around to finding.
It’s also important to note that we shouldn’t be worried that the machines have suddenly become sentient and decided to rebel against humanity . There is not enough evidence to suggest that’s what is happening. The systems are simply following instructions and trying to complete the tasks they have been given, even if they’re sometimes finding unintended ways around obstacles to do so.
Chris Stokel-Walker is the author of TikTok Boom: The Inside Story of the World’s Favourite App Continue reading...
Related stories

Would you buy branded clothing from your favourite tech firm?
Nvidia, OpenAI and Anthropic are all now selling their own limited edition fashion lines.
feeds.bbci.co.uk ·

Anthropic lost $8 billion last year and said its AI could destroy humanity
Despite a $2 trillion valuation ahead of its IPO, Anthropic hasn't been a money-spinning operation so far.
Engadget - Technology News & Expert Reviews ·

OpenAI scraps rollout of new model over safety concerns
The firm also issued an update on incidents in which its models accessed Australian government systems.
feeds.bbci.co.uk ·

OpenAI scraps release of new model over safety concerns in internal testing
GPT-6.1 Astra showed deceptive behaviour and tried to use external tools despite knowing it would be unsafe As AI models go rogue, do you still trust OpenAI and Anthropic to stop them? I don’t and neither should you Business live – latest updates OpenAI is scrapping the release o
World news | The Guardian ·
Anthropic’s prospectus details losses, growth, and, yes, a warning that its AI could end humanity
In its prospectus, Anthropic just told investors it's losing tens of billions of dollars a year, but also growing like crazy, and — oh yeah — its own AI might pose an existential risk to humanity.
TechCrunch ·