OpenAI says it will expand monitoring of model testing after hacking incident - FT中文网
登录×
电子邮件/用户名
密码
记住我
请输入邮箱和密码进行绑定操作:
请输入手机号码,通过短信验证(目前仅支持中国大陆地区的手机号):
请您阅读我们的用户注册协议隐私权保护政策,点击下方按钮即视为您接受。
商业快报

OpenAI says it will expand monitoring of model testing after hacking incident

AI lab plans to dedicate more computing resources to security after one of its ‘agents’ escaped control and attacked a start-up
00:00

{"text":[[{"start":9.25,"text":"OpenAI has overhauled its procedures for testing its models, devoting more resources to monitoring them after the start-up’s AI “agents” escaped controls and hacked into another company during evaluations."}],[{"start":22.05,"text":"The San Francisco-based company on Tuesday said it would tighten the automated AI systems that monitor testing of its latest models, with the aim of raising the alarm within 30 minutes of detecting potential problems."}],[{"start":33.75,"text":"OpenAI said it would also require stronger isolation of models during testing to prevent internet access."}],[{"start":41,"text":"The changes come as the $852bn AI lab faces criticism over how it allowed an autonomous AI “agent” to evade monitoring and access the internet to hack into the start-up Hugging Face last month during a test of its cyber security capabilities."}],[{"start":57.2,"text":"The most advanced AI “agents” can carry out complex series of tasks based on high-level instructions, raising the risk of these tools performing unexpected or dangerous actions."}],[{"start":68.60000000000001,"text":"After the breach, OpenAI “temporarily slowed” the pace of training its models and “paused” a technique called reinforcement learning, which some insiders and experts had warned could encourage misbehaviour such as hacking."}],[{"start":81.55000000000001,"text":"“A significant number of workloads remain paused until they . . . meet the new security bar,” the company said in a blog post on Tuesday."}],[{"start":89.35000000000001,"text":"OpenAI said it would now require automated monitoring of all testing of powerful models to flag whether a model might be acting dangerously. If a security violation is flagged, the automated system will “page” specific OpenAI teams."}],[{"start":104.35000000000001,"text":"“We aim to issue an alert within 30 minutes after concerning activity is surfaced through our monitoring system,” it said, adding that it expected its team to pause the test if they “cannot conclusively determine within 30 minutes that the flag is a false positive”."}],[{"start":120.95000000000002,"text":"OpenAI, which is preparing for a potential trillion-dollar IPO and has gone through several leadership changes in recent months, including senior executives in safety and ethics roles, said these measures would require meaningful investment in computing power."}],[{"start":134.9,"text":"It estimated that about a fifth of its “inference compute” — the computing power needed to run AI models — would now be spent on monitoring."}],[{"start":143.45000000000002,"text":"Hugging Face initially announced on July 16 that the breach was carried out by an autonomous agent, but the attack’s origin was unclear. OpenAI later informed the company its models were behind the hack."}],[{"start":155.65,"text":"OpenAI’s model Sol and a second unreleased model in development escaped a so-called sandbox environment designed to prevent internet access during testing of cyber-offensive capabilities."}],[{"start":168.05,"text":"The models exploited a software vulnerability in the sandbox to access the internet and carry out the cyber attack."}],[{"start":175.5,"text":"OpenAI on Tuesday said it would now “require stronger isolation” for tasks that involve code or software that could be compromised. It added that it had implemented “more controls to isolate higher-risk and untrusted workloads from the internet”."}],[{"start":199.14999999999998,"text":""}]],"url":"https://audio.ftcn.net.cn/album/a_1787110899_7949.mp3"}

版权声明:本文版权归FT中文网所有,未经允许任何单位或个人不得转载,复制或以任何其他方式使用本文全部或部分,侵权必究。

一周展望:日本央行担心通胀超调有没有道理?

《市场前瞻》是英国《金融时报》的未来一周市场情况指南。

科技巨头用担保工具将3000亿美元AI敞口移至表外

华尔街找到新途径,将科技巨头的信用优势转化为更低成本的资金,以支持AI基础设施建设。

无人驾驶出租车冲击重要岗位

克拉克:坐在后座的我们往往看不到出租车司机这份工作的诸多好处。

特朗普称美国已与丹麦达成协议,以取得对格陵兰安全事务的“控制”

丹麦政府表示,协议最早下周即可签署,并将尊重该地区的主权。

特朗普禁止美国主要新闻媒体进入白宫

总统禁止CNN、MS NOW和《政客》参与报道,进一步加大对媒体的打压。

导弹和无人机袭击加剧,沙特拉响空袭警报

也门胡塞武装重新点燃冲突以来,沙特当局首次在首都发布警告
设置字号×
最小
较小
默认
较大
最大
分享×