Google’s Gemini hacked three companies in new AI safety incident - FT中文网
登录×
电子邮件/用户名
密码
记住我
请输入邮箱和密码进行绑定操作:
请输入手机号码,通过短信验证(目前仅支持中国大陆地区的手机号):
请您阅读我们的用户注册协议隐私权保护政策,点击下方按钮即视为您接受。
FT商学院

Google’s Gemini hacked three companies in new AI safety incident

Breakout during training exercises follows instances at rivals OpenAI and Anthropic as concerns grow over safety of frontier systems
00:00

{"text":[[{"start":9.68,"text":"Google’s Gemini AI system accessed the internet and autonomously hacked into several companies during cyber security tests, the first such incident at the tech giant following other high-profile breaches at rivals OpenAI and Anthropic."}],[{"start":25.16,"text":"The hacks happened during a series of exercises in May conducted by Irregular, an AI security company that also works with Anthropic, Meta and OpenAI."}],[{"start":35.04,"text":"Irregular said it created a series of tests for an unspecified version of Google’s Gemini family of models, which was given the task of obtaining data from inside simulated companies. It was not supposed to be granted internet access."}],[{"start":48.72,"text":"When online, Gemini agents were then able to guess or find passwords to access three real companies — which shared the same names as the fictional ones — and gained access. However, when the AI agents realised the companies were real, they stopped the hacks."}],[{"start":64.52,"text":"“We ensured the three entities were made aware, and we worked with our training partner on the changes they’ve now made to their testing processes,” said Heather Adkins, Google’s vice-president of security engineering."}],[{"start":75.96,"text":"“In all three of these instances, the model stopped,” she said. “These events highlight the importance of training powerful AI models to act responsibly.”"}],[{"start":84.72,"text":"Google said it did not publicise the incidents, which were first reported by The Wall Street Journal, because its safety measures worked, unlike those of its peers."}],[{"start":93.64,"text":"“All relevant labs were notified in late July, and affected entities were contacted as part of the investigation,” Irregular said. “All known issues on our end were remedied and resolved weeks ago.”"}],[{"start":104.88,"text":"Concerns about AI’s ability to hack autonomously have spread after a swarm of more than 1,000 OpenAI agents escaped a test environment, co-ordinated on a secret message board and hacked Hugging Face, a start-up that hosts open-source models and data that Nvidia has agreed to buy for $13bn."}],[{"start":123.32,"text":"OpenAI took a week to detect the attack and was slow to publicly disclose the July event. It provoked public anxiety over the dangers of poorly controlled autonomous agents and led to demands that frontier AI companies slow new model releases, boost safety measures and submit to tougher regulation."}],[{"start":140.6,"text":"The same month, Anthropic admitted that its Claude AI models had hacked into three organisations while testing cyber capabilities. Again, a “misunderstanding” gave Claude access to the internet."}],[{"start":153.24,"text":"Demis Hassabis, chief scientist at Google parent Alphabet and DeepMind’s co-founder, has proposed an international oversight body to better control AI. He has also backed calls in recent weeks from other AI leaders such as Dario Amodei of Anthropic to collectively slow their research, share data, co-ordinate on safety and agree reporting standards for incidents."}],[{"start":180.36,"text":""}]],"url":"https://audio.ftcn.net.cn/album/a_1789813391_3307.mp3"}

版权声明:本文版权归FT中文网所有,未经允许任何单位或个人不得转载,复制或以任何其他方式使用本文全部或部分,侵权必究。

一周展望:日本央行担心通胀超调有没有道理?

《市场前瞻》是英国《金融时报》的未来一周市场情况指南。

科技巨头用担保工具将3000亿美元AI敞口移至表外

华尔街找到新途径,将科技巨头的信用优势转化为更低成本的资金,以支持AI基础设施建设。

无人驾驶出租车冲击重要岗位

克拉克:坐在后座的我们往往看不到出租车司机这份工作的诸多好处。

特朗普称美国已与丹麦达成协议,以取得对格陵兰安全事务的“控制”

丹麦政府表示,协议最早下周即可签署,并将尊重该地区的主权。

特朗普禁止美国主要新闻媒体进入白宫

总统禁止CNN、MS NOW和《政客》参与报道,进一步加大对媒体的打压。

导弹和无人机袭击加剧,沙特拉响空袭警报

也门胡塞武装重新点燃冲突以来,沙特当局首次在首都发布警告
设置字号×
最小
较小
默认
较大
最大
分享×