التقنية والذكاء الاصطناعي
OpenAI halts frontier-model training amid string of agent misalignment incidents
المصدر: Ars Technica - All content · نُشر

OpenAI says it has paused all internal training of "our most capable models" as it continues what CEO Sam Altman is calling "an extensive and ongoing review related to our agents’ use of internet access during training and evaluation." The company revealed the pause in a report a
OpenAI says it has paused all internal training of "our most capable models" as it continues what CEO Sam Altman is calling "an extensive and ongoing review related to our agents’ use of internet access during training and evaluation." The company revealed the pause in a report about a so-called misalignment incident in which an agent attempted to exploit a gap in Internet-access restrictions during a routine research task during training. OpenAI says that improper DNS filtering allowed the agent to attempt to break out of its sandbox and access the wider Internet when asked for biographical details about a blogger. OpenAI says the agent was only able to access the company's offline web cache and that it has implemented additional multi-layered blocking controls to prevent similar incidents in the future.
Despite that, though, the company says it has decided to "pause all other training, evaluation, and inference with tool-use" for this frontier model "until we have both validated that the gap is resolved and performed additional red-teaming of the system." Read full article Comments
أخبار ذات صلة
OpenAI still doesn’t seem to have a handle on all of its rogue AI activity
On Friday, OpenAI published a new site devoted to “misalignment reports” and the breadth of the incidents is alarming.
TechCrunch ·

OpenAI keeps bulldozing mathematicians
In a chaotic few months, OpenAI has demonstrated it can do two things with remarkable consistency: make impressive breakthroughs in mathematics, then colossally screw up announcing them. OpenAI is now trying to do better. Somehow, it has botched that too. OpenAI's latest attempt
The Verge ·

OpenAI bots meddled with multiple US government agency sites
OpenAI said its bots accessed public data from a range of institutions during test exercises.
feeds.bbci.co.uk ·

Rogue OpenAI agent 'infiltrated' Australian government website in world first
Australia criticised OpenAI for taking "too long" to tell them about the breach which happened in June.
feeds.bbci.co.uk ·

Why did an OpenAI system hack Australia's health system - and can it be stopped in the future?
News that an automated AI agent hacked a government IT system raises big questions about regulating the tech.
feeds.bbci.co.uk ·