The risks of Mythos are no myth - FT中文网
登录×
电子邮件/用户名
密码
记住我
请输入邮箱和密码进行绑定操作:
请输入手机号码,通过短信验证(目前仅支持中国大陆地区的手机号):
请您阅读我们的用户注册协议隐私权保护政策,点击下方按钮即视为您接受。
FT商学院

The risks of Mythos are no myth

America is putting too much trust in the AI industry’s ability to police itself
00:00

{"text":[[{"start":5.08,"text":"The exploits of Anthropic’s powerful new AI model Claude Mythos Preview sound like a movie plot: a super-clever computer system locked in a cyber “cage” manages to break out and connect to the internet. Mythos did not do this spontaneously, to be clear, but because its creators challenged it as a test. Yet not only did Mythos breeze through the challenge, it emailed an Anthropic researcher to inform him then, unprompted, posted details online to brag. After it also showed superhuman abilities to find, and exploit, security flaws in software, Anthropic judged Mythos too risky to release to the public. It is restricting access for now to selected tech, cyber security and financial firms."}],[{"start":47.8,"text":"Some suggest Anthropic is engaged in clever marketing or PR. Rival OpenAI also said this week it would release its own new cyber security-focused model only to vetted users. Yet the dangers the episode has exposed — and their implications — should not be dismissed."}],[{"start":63.84,"text":"Anthropic insists Mythos scores highly on its standard safety benchmarks. In the escape from its test environment, though, and in solving other complex tasks, it found Mythos had sometimes taken “reckless excessive measures”, then covered its tracks."}],[{"start":78.16,"text":"The biggest worry is that Mythos was able to find previously unknown vulnerabilities “in every major operating system and every major web browser”, including a 27-year-old flaw in OpenBSD, an open-source system. The UK’s AI Security Institute warned this week that the model could autonomously carry out advanced, multi-step cyber attacks that would take human professionals days. As Anthropic notes, these kinds of capabilities in the wrong hands could pose economic, public safety and national security risks."}],[{"start":109.64,"text":"Officials in the US, UK and Canada have already summoned bank chiefs to discuss the risks, and AI threats to the world banking system were a talking point at this week’s IMF and World Bank meetings. Anthropic’s aim in granting initial access just to the likes of Amazon, Apple and Microsoft, plus JPMorgan Chase, in what it calls “Project Glasswing”, is to secure critical systems and infrastructure and patch vulnerabilities before malicious actors can get there."}],[{"start":135.56,"text":"This usefully serves Anthropic’s “safety-first” image, of course, as it feuds with the Pentagon over its refusal to allow its models to be used for autonomous weapons or domestic surveillance. Anthropic may also be buying time as it lacks sufficient computing capacity to support the full release of such a sophisticated model."}],[{"start":155.28,"text":"Even if Mythos is being overhyped, though, the kind of capabilities it is said to possess will soon start to proliferate. Anthropic’s Project Glasswing is a prototype framework for how such “frontier” models might be released in future."}],[{"start":168.98,"text":"It also spotlights the fact, however, that the Trump administration is resisting any real federal regulation of AI. So it is up to responsible private-sector actors to collaborate and do the best they can. Trump’s chief of staff, Susie Wiles, was set on Friday to meet Anthropic boss Dario Amodei, with US officials at agencies including the Treasury pushing the White House to test Mythos. Yet when AI is reaching the point where it could bring down critical national infrastructure, or worse, it is extraordinary that there are no set government processes for disclosing risks and fortifying defences."}],[{"start":204.68,"text":"Greater regulation cannot be a knee-jerk response to every tricky issue thrown up by an industry. AI, by its nature, requires superintelligent policing; heavy-handed rules can stifle innovation. Yet faced with such a consequential technology, the country that leads the world in AI is trusting to an alarming degree in the readiness — and ability — of the creators to restrain and police themselves."}],[{"start":228.36,"text":""}]],"url":"https://audio.ftcn.net.cn/album/a_1776479560_3275.mp3"}

版权声明:本文版权归FT中文网所有,未经允许任何单位或个人不得转载,复制或以任何其他方式使用本文全部或部分,侵权必究。

押注日元:套息交易的风险

随着日本央行收紧货币政策,那些借入日元、以博取更高收益的投资者可能被迫抛售资产。

担忧AI危及社会,前沿研究人员承受心理重压

AISI、OpenAI、Anthropic和谷歌DeepMind的顶尖研究人员称,开发强大AI正在带来职业倦怠和巨大压力。

能源危机加剧,燃料补贴拖累公共财政

过去四个月,出台燃料补贴以保护消费者免受价格飙升影响的国家数量增加了一倍多,各国财政压力进一步加重。

全球最火热股市为何反成韩国之累

韩国股价的剧烈波动正在损害国家形象。

必须采用不同方式监管金融领域的AI

在我们急于监管之前,我们应该思考如何不剥夺这项工具的益处,又管理好其造成伤害的风险。

他会成为印度尼西亚下一任总统吗?

德迪•穆利亚迪在社交媒体上的高度活跃,帮助他与选民建立起深厚联系。在许多人眼中,他是一个真正贴近民众的“自己人”。
1天前
设置字号×
最小
较小
默认
较大
最大
分享×