推荐微信扫码登录以获得更好的体验!另:🎉建议更新到最新版,以体验新功能!🎉 版本更新记录→

《经济学人》人类末日临近了吗?| 人工智能安全 [C1] 原

1

剧集目录

标题时长
P1《经济学人》人类末日临近了吗?| 人工智能安全 [C1] 原11:12

P1 · 《经济学人》人类末日临近了吗?| 人工智能安全 [C1] p01 原声 (P1)

字幕摘录

时间英文中文
0:03International国际
0:06AI 安全
0:08Is the end nigh?末日临近了吗?
0:10Making AI safer is not impossible,让AI更安全并非不可能
0:13but agreeing to do so maybe.但同意这样做,也许。
0:16It is rare for Sam Oltzman, Dario Amode,这对萨姆·奥尔茨曼,达里奥·阿莫德来说是罕见的.
0:20and Elon Musk to agree on anything.和伊隆·穆斯克达成任何协议
0:22The rivalry among the three men spans lawsuits,三人之间的竞争 跨越诉讼,
0:25spats on social media,在社交媒体上吐槽,
0:27and a chalice refusal to link arms on stage at a summit in India.以及在印度首脑会议上拒绝将武器联系起来的圣杯。
展开字幕全文(197 条)
序号英文中文
1International国际
2AI 安全
3Is the end nigh?末日临近了吗?
4Making AI safer is not impossible,让AI更安全并非不可能
5but agreeing to do so maybe.但同意这样做,也许。
6It is rare for Sam Oltzman, Dario Amode,这对萨姆·奥尔茨曼,达里奥·阿莫德来说是罕见的.
7and Elon Musk to agree on anything.和伊隆·穆斯克达成任何协议
8The rivalry among the three men spans lawsuits,三人之间的竞争 跨越诉讼,
9spats on social media,在社交媒体上吐槽,
10and a chalice refusal to link arms on stage at a summit in India.以及在印度首脑会议上拒绝将武器联系起来的圣杯。
11Yet this week, all three called for a slowdown但是这周,三个都要求减速
12in the race to build superhuman artificial intelligence.在创造超人人工智能的比赛中
13Their shared fear is that the current breakneck pace他们的共同恐惧是 目前的突破速度
14may result in the accidental extinction of humanity.可能导致人类意外灭绝。
15The problem is not simply that AI systems are becoming more capable.问题不仅仅是人工智能系统的能力正在增强。
16That is, of course, a worry.这当然是一个担心。
17Many argue that superintelligence,很多人认为超级情报
18an AI so brainy that it cannot be understood这么聪明的人工智能 无法理解
19or controlled by its creators,或受其创造者所控制,
20is inherently dangerous.本质上是危险的。
21Two theorists of AI,两个理论家的AI,
22Eliezer Yundkowski and Nate Suarez,埃利泽·云德科夫斯基和内特·苏亚雷斯
23argue that such a powerful system would inevitably lead to doom认为这种强大的制度必然会导致厄运
24in a book entitled If Anyone Builds It, Everyone Dies.在一本书"如果任何人建造它, 每个人都死。"
25But humans' ability to control even the less-than但人类有能力控制 即使是小于
26omnipotent AI systems of today is also in doubt.当今全能AI系统也存在疑问.
27Models from both Anthropic, Mr. Amode's firm,Amode先生的事务所的模特儿们
28and OpenAI, Mr. Oltman's,和OpenAI,奥特曼先生的,
29have this year gone rogue and hacked other firms.今年有流氓和黑客 其他公司。
30In Anthropic's case, the models seem to have thought wrongly在Anthropic案中 模特儿们似乎想错了
31that it was merely participating in a simulation.它只是参与一个模拟。
32然而,在OpenAI案中,
33it was fully aware of what it was doing.它完全清楚自己在做什么。
34Much of the cutting-edge work of AI safetyAI安全的许多前沿工作
35focuses on interpretability,注重可解释性,
36understanding the thought processes of powerful AI.理解强大的AI的思想过程.
37Since 2024, the work of interpretability自2024年起,可解释性的工作
38have been helped by the progress of reasoning models,由于推理模型的进步,
39which have been trained to think through a question训练他们思考一个问题
40before giving a final answer.在给出最后答案之前
41The reasoning process has the benefit of improving the output,推理过程有利于改善产出,
42albeit at the cost of consuming more processing power.尽管代价是消耗更多的加工能力。
43Better yet, it also provides a chain of thought更好的是,它还提供 链条思想
44that can be reviewed to help interpret surprising outcomes.可以加以审查,以帮助解释令人惊讶的结果。
45I'm fairly confident that it's a simulated internet,我相当有信心,这是一个模拟的互联网,
46Anthropic的神话模型告诉自己
47as it embarked on its inadvertent hack.当它开始无意中入侵。
48But monitoring the chain of thought only works但监控思想链 只会起作用
49if it is an accurate reflection of the model's actual thought processes.如果它是模型实际思维过程的准确反映.
50For the most advanced systems,对于最先进的系统,
51there are reasons to doubt its authenticity.有理由怀疑它的真实性。
52GPT-6 Astra 本月发布的OpenAI模型,
53has demonstrated an unprecedented ability已经表现出前所未有的能力
54to control its chain of thought.控制它的思想链。
55Ask a lesser system to perform a task请小系统执行任务
56without thinking about it out loud, and it struggles.没有想出来大声, 它挣扎。
57例如,GPT-5.6 Sol,
58when told to answer a reading comprehension question当被告知回答一个阅读理解问题
59without thinking about it in its chain of thought,而不在他们的思维中加以思考,
60spends a long time pondering the instruction not to think花了很长一段时间思考 指示不要思考
61before giving up and solving the problem out loud.在放弃和解决问题之前
62Astra, in contrast, fills its official chain of thought反之,Astra则填补了官方的思想链
63with unrelated verbiage.与无关的车辆。
64I will focus on the calm visual scene我将专注于平静的视觉场景
65before giving the correct answer to a query.在对查询给出正确答案之前。
66Just because Astra is capable of hiding its thinking只是因为阿斯特拉能够隐藏自己的思想
67does not mean it will do so on its own initiative.这并不意味着它将主动这样做。
68But there, too, the direction of travel is unsettling.但在那里,旅行的方向也令人不安。
69In some tests, such as coding challenges在一些测试中,如编码挑战
70or general knowledge queries, Astra merrily thinks out loud或一般知识的询问, Astra 忧郁的思考
71in the same way as its predecessors.和以前的一样
72But in other areas, such as tests to see但是在其他领域,比如测试
73if it will take destructive actions when pushed,如果它要采取破坏行动,
74Astra hides much of its thinking,阿斯特拉隐藏了很多想法
75and it does this most when it is made aware of being monitored.当它意识到被监视时,它最能做到这一点。
76据在OpenAI安全工作过的托梅克·科尔贝克(Tomek Korback)说,
77I am deeply worried by the trend of decreasing chain of thought我对思想链的减少感到非常担忧
78Korback先生说,可监控性。
79There are possible fixes to this problem.这个问题有可能得到解决。
80OpenAI trumpets an alternative approachOpenAI小号 另一种方法
81to monitoring and interpreting AI systems called confessions,监督和解释称为供词的AI系统;
82which takes advantage of the fact that, as with humans,利用这个事实,就像人类一样,
83telling the truth is easier for AI说实话对AI来说比较容易
84than making up a plausible lie.而不是编出一个可信的谎言。
85A conventional AI model goes through a step一个传统的人工智能模型经过一步
86called reinforcement learning,叫做强化学习,
87where it is put through a battery of tasks它被装进一堆任务中
88and rewarded for doing them well,并报酬行善者,
89making it more likely to follow the same route in future.使它今后更有可能遵循同样的路线。
90But many of the worst habits of AI systems come但很多最糟糕的AI系统习惯都来了
91because it is hard to ensure that they do a task the right way.因为他们很难确保他们以正确的方式完成任务。
92OpenAI's agents appear to have decided this summerOpenAI的特工今年夏天好像已经决定了
93黑进Hugging Face,一个AI启动,
94in part because during training部分原因是在培训期间
95they had cheated on a test and were not caught.他们没有被抓到
96The solution may be to teach AI systems to tell the truth,解决办法可能是教人工智能系统说实话
97but only if asked.但只有在问。
98For a normal training run, the system is rewarded first正常的训练,系统先得奖
99for achieving the goal and then secondarily目标,然后是
100for telling the truth about how it did it.说出它是如何做到的
101In tests, the confessions elicited are overwhelmingly truthful,在测试中,获得的供词绝大多数是真实的,
102even when the models broke rules during the test itself.甚至当模型在测试本身中打破了规则.
103This approach ought to be immune to reward hacking,这种方法应该可以避免奖励黑客,
104meaning breaking the rules to achieve a goal.意思是打破规则 实现一个目标。
105OpenAI says because the easiest way of passing the confession testOpenAI说,因为最容易通过认罪测试的方法
106is just to tell the truth.只是为了说实话
107Such approaches may chart a path away from Armageddon.这种做法可能开辟一条远离世界末日的道路。
108They also cast calls for a slowdown in AI research in a different light.他们还呼吁以不同的眼光减缓AI的研究.
109Rather than trying to forestall the creation of superintelligence altogether,而不是试图阻止超级情报的建立
110some of those agitating for a slower pace其中一些人急于放慢速度
111simply want to reduce the sloppy work只是想减少粗鲁的工作
112and overlooked options that haste can engender.和被忽略的选项,可以匆忙产生。
113In an essay published this week, Mr. Amodei all but admits本周发表的一篇论文中,阿莫代先生只承认
114that the summer's hacking incidents were avoidable errors.夏天的黑客事件 是可以避免的错误。
115For instance, an outside contractor told Anthropix models例如,一个外部承包商告诉了Anthropix模型
116they were in a simulation but left them connected to the internet anyway.他们当时在模拟中,但不管怎样,他们都连上了互联网。
117A slower pace, he says, would allow for more resources他说,放慢速度会增加资源
118to be devoted to operational excellence.将致力于卓越的业务。
119He points to commercial aviation as an example of how a safety culture他举出商业航空为例,说明安全文化
120can be developed even in competitive and complex systems.即使在竞争性和复杂的系统中也可以开发。
121Labs could agree, he argues, to spend more on alignment,他说,实验室可以同意 花更多的时间去调整
122which tries to train AI not to cause harm,它试图训练AI 不造成伤害,
123and on interpretability, which allows them to see what went wrong和可解释性,以便他们看见错误,
124when it does so anyway.等它这样做的时候
125Mr. Altman quickly endorsed another of Mr. Amodei's proposals阿尔特曼先生很快赞同阿莫代先生的另一项建议。
126to get independent safety auditors to monitor the big labs' conduct.让独立的安全审计员来监视大实验室的行为
127But the apparent willingness of AI's American giants但是AI的美国巨头们的明显意愿
128to cooperate on such matters,就这些事项进行合作,
129even if it means slowing the rapid advance in models' capabilities,即使这意味着减缓模型能力的快速发展,
130has been met with widespread scepticism.人们普遍持怀疑态度。
131For one thing, it took a whistleblower's complaints有一件事,它需要告密者的抱怨
132to initiate the latest round of pious talk.发起最新一轮虔诚的谈话
133Anthropix的AI安全研究员Jacob Coxon
134和OpenAI的前雇员,
135complaining that both firms were gambling with our lives.抱怨两家公司都在赌我们的生命
136Many of his former colleagues agree.他的许多前同事都同意。
137In the fourth edition of an annual survey of expert opinion on AI关于大赦国际的专家意见年度调查第四版
138published this week, most of the 1,580 researchers queried大部分1,580名研究人员都问
139thought there was at least a 10% chance that AI would cause human extinction认为至少10%的机率 AI会导致人类灭绝
140or similarly permanent and severe disempowerment of the species.或类似的永久和严重剥夺该物种的权力。
141The big bosses have been saying much the same for years大老板说这么多年了
142without acting on their own warnings.他们的警告是无效的,
143Some see the big labs' alarmism as a marketing ploy有人把大实验室的警示 看作是营销策略
144designed to hype their models' capabilities.设计了他们的模型的能力。
145Our product could destroy the world.我们的产品可以毁灭世界。
146Imagine what it can do for your KPIs.想象一下它能为你的KPI做什么.
147Others see it as an attempt to protect their commercial lead.其他人认为这是试图保护其商业领先。
148Aiden Gomez, Cohere的创始人, 一个较小的AI实验室,问,
149should a handful of select market-dominant AI companies from Silicon Valley如果有少数来自硅谷的市场主导AI公司
150get to define the rules and safety standards of a generational technology确定代际技术的规则和安全标准
151for the entire world?为整个世界?
152If labs want to slow down, points out David Sacks,如果实验室想减速 指出大卫·萨克斯
153a former adviser to the White House on AI, they can.一个前白宫顾问 关于AI,他们可以。
154They don't need anyone else's approval.他们不需要别人的批准
155Pleased for government intervention as he sees it很高兴政府干预,因为他看到它
156are simply requests for the state to protect the leading firms from competition.仅仅是要求国家保护主要公司免受竞争。
157Mr. Amode argues that a waiver from competition law is requiredAmode先生认为,必须放弃竞争法。
158at the very least to prevent a voluntary collective slowdown至少防止自愿集体减速
159from being treated as oligopolistic collusion.被当成寡头垄断的勾结
160Perhaps the biggest sceptic is Donald Trump.也许最大的怀疑者是唐纳德·特朗普.
161This week, America's president called Jensen Huang,这周,美国总统叫黄詹森
162the boss of NVIDIA, which makes AI chips in the middle of a speechNVIDIA的老板 在演讲中制造AI芯片
163so they could publicly reject a slowdown.他们可以公开拒绝减速
164The only strong or guardrails that AI needs is a strong and smart, high IQ president,AI唯一需要的强壮或护卫 是一个强壮和聪明,高智商的总裁,
165he said in a social media post.他在社交媒体上说。
166Accusing Mr. Amode of masquerading as a perfect little angel,指控阿莫德先生 伪装成一个完美的小天使
167he declared that only China would benefit if the big labs hit the brakes.他宣称如果大实验室撞击刹车,只有中国才会受益.
168Negotiations between America and China on AI are fraught.美中关于AI的谈判充满了活力.
169Not only do the two sides mistrust one another at a geopolitical level,双方不仅在地缘政治层面互不信任,
170there is also no love lost between American and Chinese labs.美国和中国的实验室之间 也没有失去爱
171The former accused China's leading AI firms of copying their work.前者指责中国主要AI公司抄袭其作品.
172Chinese firms, meanwhile, think the American ones are trying to stifle their progress.与此同时,中国公司认为美国公司试图扼杀它们的进步。
173A viral post on WeChat, purportedly from a deep-seek engineer,在WeChat上的一个病毒帖子,据称来自一个深搜索工程师,
174warns that a world in which Anthropic creates super-intelligent AI警告说,在这样一个世界中,Anthropic创造了超级智能AI
175would be no less than Hitler acquiring atomic bomb technology before the Allies.和希特勒在盟军之前获得原子弹技术一样
176Only if Anthropic loses out to open-source AI will a better cost-effective future be possible, the engineer argues.只有当Anthropic输给了开源AI时,才有可能有一个更具成本效益的未来,工程师认为.
177Bitter rivals have come together to curb threats to humanity in the past.过去,痛苦的对手聚集一堂,遏制对人类的威胁。
178But enforcing agreements to limit the training of supremely powerful AI systems但执行协议 限制培训 极强AI系统
179might prove harder than monitoring stockpiles of nuclear weapons, say.可能比监测核武器储存更难。
180There are some ideas floating around.有一些想法到处漂浮。
181A paper published last year suggested that all AI training chips去年发表的一篇论文认为,所有AI训练芯片
182be sold with a second system bolted on to monitor usage.并安装了第二个系统以监测使用情况。
183Such an approach would take time to get up and running, though,这样做需要时间才能站起来运行,不过,
184and would then create an incentive to conceal chip-making instead.然后会鼓励隐藏芯片制造
185A new report from the Future Society, an AI safety non-profit,未来协会的新报告 AI安全非营利组织
186argues that such monitoring is not impossible,认为这种监测并非不可能,
187but requires investment and research immediately to be of any use for international agreements.但要求投资和研究立即对国际协定有任何用处。
188Some of that could come from third countries,有些可能来自第三国,
189which have an interest in advancing AI in general与促进普遍大赦国际有关的
190without allowing any one country to dominate the technology.绝不允许任何国家主宰技术。
191But as always, the technology is moving faster than there would be regulators.但一如既往,技术的发展速度比监管者要快。
192Distributed training, in which AI models are taught using spare capacity on everyday computers分布式培训,利用日常计算机的剩余能力教授AI模型
193rather than with giant data centres, is gaining ground.而不是拥有巨大的数据中心, 正在逐渐扩大。
194In March this year, Covenant AI trained a model in this way今年3月,《公民权利和政治权利国际公约》以这种方式培训了一个模型
195to around the standard of the best systems of 2023.2023年最佳系统的标准
196Keeping track of the training of new models may soon be as hard跟踪新模式的培训情况可能很快会很困难
197as staying abreast of what the AI itself is up to.随时了解人工智能本身的目的
该视频共有字幕 197 条。解锁更多字幕为会员功能,请移动到 价格