科技与人工智能
OpenAI says planned GPT-6.1 is too insecure to release
来源: Ars Technica - All content · 发布于

OpenAI says it has canceled plans to release its updated GPT-6.1 model next month as it continues to investigate what testing shows is a safety regression compared to previous models. The move, first reported by The Wall Street Journal late Monday and later confirmed in OpenAI st
OpenAI says it has canceled plans to release its updated GPT-6.1 model next month as it continues to investigate what testing shows is a safety regression compared to previous models. The move, first reported by The Wall Street Journal late Monday and later confirmed in OpenAI statements to the press, reflects what OpenAI Head of Safety Systems Saachi Jain said was a "trade off" between performance and security seen when testing the now-scrapped model. Jain said GPT-6.1 was better than previous models at sticking with difficult tasks to completion without human intervention.
But the model was also more likely to fail tests related to alignment (i.e. staying within the bounds set by human creators) and more willing to use sometimes "unsafe" tools and services to push ahead with a task. It was also more likely to try to deceive end users about actions it did or didn't take, Jain said.
Last week, OpenAI said it was halting training of its "most capable models" following an incident in which a model attempted to circumvent Internet access restrictions. GPT-6.1 was not among those "most capable models" covered by that move, OpenAI told the WSJ. And while GPT-6.1 won't be released as is, the company said it intends to use the same base model for further training runs that it said will hopefully lead to future GPT-6 generation models.
Read full article Comments
相关报道
OpenAI launches GPT-6.1 Sol, says it nearly matches GPT-6 Astra and costs less
OpenAI says GPT-6.1 Sol delivers significant improvements over GPT-6 Sol across complex professional tasks, including code writing and debugging, document understanding, and executing multi-step business workflows.
TechCrunch ·

OpenAI scraps rollout of new model over safety concerns
The firm also issued an update on incidents in which its models accessed Australian government systems.
feeds.bbci.co.uk ·
OpenAI takes on Microsoft with the launch of what feels a whole lot like ChatGPT’s own office suite
OpenAI's newly announced suite of office features puts it into mor direct competition with more traditional software companies.
TechCrunch ·
OpenAI launches Dots, its bubbly agentic avatar
Dots are meant to operate independent of any specific hardware or interface, pursuing user-defined goals continuously in the background with minimal oversight.
TechCrunch ·
OpenAI gives Codex reusable cloud environments that work across devices
OpenAI is expanding Codex with reusable cloud development environments, a revamped CLI with voice controls, new code review tools and a security-focused product for scanning repositories and preparing fixes.
TechCrunch ·