Artificial Intelligence, Values, and Alignment
Iason Gabriel
Google DeepMind (United Kingdom)
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
This paper looks at philosophical questions that arise in the context of AI alignment. It defends three propositions. First, normative and technical aspects of the AI alignment problem are interrelated, creating space for productive engagement between people working in both domains. Second, it is important to be clear about the goal of alignment. There are significant differences between AI that aligns with instructions, intentions, revealed preferences, ideal preferences, interests and values. A principle-based approach to AI alignment, which combines these elements in a systematic way, has considerable advantages in this context. Third, the central challenge for theorists is not to identify ‘true’ moral principles for AI; rather, it is to identify fair principles for alignment that receive reflective endorsement despite widespread variation in people’s moral beliefs. The final part of the paper explores three ways in which fair principles for AI alignment could potentially be identified.
逐年被引趋势
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
可就本文提问;依据不足时会说明。
学术脉络
学科主题
社会科学Ethics and Social Impacts of AI
Explainable Artificial Intelligence (XAI) · Artificial Intelligence in Healthcare and Education
参考文献 85
此处列出前 3 条
引用本文 708
按被引量排序,此处列出前 3 条