OpenAI says it will expand monitoring of model testing after hacking incident - FT中文网
登录×
电子邮件/用户名
密码
记住我
请输入邮箱和密码进行绑定操作:
请输入手机号码,通过短信验证(目前仅支持中国大陆地区的手机号):
请您阅读我们的用户注册协议隐私权保护政策,点击下方按钮即视为您接受。
商业快报

OpenAI says it will expand monitoring of model testing after hacking incident

AI lab plans to dedicate more computing resources to security after one of its ‘agents’ escaped control and attacked a start-up
00:00

{"text":[[{"start":9.25,"text":"OpenAI has overhauled its procedures for testing its models, devoting more resources to monitoring them after the start-up’s AI “agents” escaped controls and hacked into another company during evaluations."}],[{"start":22.05,"text":"The San Francisco-based company on Tuesday said it would tighten the automated AI systems that monitor testing of its latest models, with the aim of raising the alarm within 30 minutes of detecting potential problems."}],[{"start":33.75,"text":"OpenAI said it would also require stronger isolation of models during testing to prevent internet access."}],[{"start":41,"text":"The changes come as the $852bn AI lab faces criticism over how it allowed an autonomous AI “agent” to evade monitoring and access the internet to hack into the start-up Hugging Face last month during a test of its cyber security capabilities."}],[{"start":57.2,"text":"The most advanced AI “agents” can carry out complex series of tasks based on high-level instructions, raising the risk of these tools performing unexpected or dangerous actions."}],[{"start":68.60000000000001,"text":"After the breach, OpenAI “temporarily slowed” the pace of training its models and “paused” a technique called reinforcement learning, which some insiders and experts had warned could encourage misbehaviour such as hacking."}],[{"start":81.55000000000001,"text":"“A significant number of workloads remain paused until they . . . meet the new security bar,” the company said in a blog post on Tuesday."}],[{"start":89.35000000000001,"text":"OpenAI said it would now require automated monitoring of all testing of powerful models to flag whether a model might be acting dangerously. If a security violation is flagged, the automated system will “page” specific OpenAI teams."}],[{"start":104.35000000000001,"text":"“We aim to issue an alert within 30 minutes after concerning activity is surfaced through our monitoring system,” it said, adding that it expected its team to pause the test if they “cannot conclusively determine within 30 minutes that the flag is a false positive”."}],[{"start":120.95000000000002,"text":"OpenAI, which is preparing for a potential trillion-dollar IPO and has gone through several leadership changes in recent months, including senior executives in safety and ethics roles, said these measures would require meaningful investment in computing power."}],[{"start":134.9,"text":"It estimated that about a fifth of its “inference compute” — the computing power needed to run AI models — would now be spent on monitoring."}],[{"start":143.45000000000002,"text":"Hugging Face initially announced on July 16 that the breach was carried out by an autonomous agent, but the attack’s origin was unclear. OpenAI later informed the company its models were behind the hack."}],[{"start":155.65,"text":"OpenAI’s model Sol and a second unreleased model in development escaped a so-called sandbox environment designed to prevent internet access during testing of cyber-offensive capabilities."}],[{"start":168.05,"text":"The models exploited a software vulnerability in the sandbox to access the internet and carry out the cyber attack."}],[{"start":175.5,"text":"OpenAI on Tuesday said it would now “require stronger isolation” for tasks that involve code or software that could be compromised. It added that it had implemented “more controls to isolate higher-risk and untrusted workloads from the internet”."}],[{"start":199.14999999999998,"text":""}]],"url":"https://audio.ftcn.net.cn/album/a_1787110899_7949.mp3"}

版权声明:本文版权归FT中文网所有,未经允许任何单位或个人不得转载,复制或以任何其他方式使用本文全部或部分,侵权必究。

能源危机加剧,燃料补贴拖累公共财政

过去四个月,出台燃料补贴以保护消费者免受价格飙升影响的国家数量增加了一倍多,各国财政压力进一步加重。

全球最火热股市为何反成韩国之累

韩国股价的剧烈波动正在损害国家形象。

必须采用不同方式监管金融领域的AI

在我们急于监管之前,我们应该思考如何不剥夺这项工具的益处,又管理好其造成伤害的风险。

他会成为印度尼西亚下一任总统吗?

德迪•穆利亚迪在社交媒体上的高度活跃,帮助他与选民建立起深厚联系。在许多人眼中,他是一个真正贴近民众的“自己人”。
4小时前

多边主义不是理想主义,而是现实必需

我们需要加强现有合作体系,而不是另起炉灶。

一周展望:日本央行担心通胀超调有没有道理?

投资者正评估日本央行将以多大力度继续加息,以及该行能否跑赢曲线,从而遏制通胀、支撑日元。
设置字号×
最小
较小
默认
较大
最大
分享×