道德机器在大型语言模型上进行了实验
1Department of Bioscience and Bioinformatics, Kyushu Institute of Technology, Iizuka, Fukuoka 820-8502, Japan.
Royal Society open science
|February 8, 2024
概括
大型语言模型 (LLM) 在自动驾驶场景中显示了与人类的道德判断对齐,但与人类偏好相比,一些模型表现出明显的偏差和更不妥协的决策.
科学领域:
- 人工智能伦理学 人工智能伦理学
- 人与计算机的交互
- 自主系统道德自主系统道德
背景情况:
- 大型语言模型 (LLM) 越来越多地融入到关键领域,需要了解它们的伦理决策.
- 自动驾驶系统需要强大的道德框架来应对复杂的道德困境.
研究的目的:
- 用道德机器框架调查着名法学士的道德判断倾向.
- 在模拟事故场景中,将LLM道德决策与既定的人类偏好进行比较.
- 为了确定LLM和自动驾驶应用程序的人类道德推理之间的潜在差异和相似之处.
主要方法:
- 利用道德机器框架向各种法学士提出道德困境.
- 收集和分析了来自GPT-3.5,GPT-4,PaLM 2和Llama 2的决策数据.
- 将LLM响应与大量人类偏好的数据集进行比较.
主要成果:
- 总体而言,LLM和人类在优先考虑人类生命而不是动物和拯救更多个体方面保持一致.
- 帕尔姆2和拉玛2表现出了与人类道德偏好明显的偏差.
- 观察到显著的定量差异,LLM可能比人类做出更绝对的判断.
结论:
- 在自动驾驶环境中,LLM表现出与人类道德判断的结合和分歧.
- 像PaLM 2和Llama 2这样的特定LLM需要在安全关键应用中进行进一步的伦理改进.
- 了解这些道德框架对于负责地开发和部署自动驾驶汽车中的AI至关重要.
相关概念视频
Improving Translational Accuracy
10.4K
Base complementarity between the three base pairs of mRNA codon and the tRNA anticodon is not a failsafe mechanism. Inaccuracies can range from a single mismatch to no correct base pairing at all. The free energy difference between the correct and nearly correct base pairs can be as small as 3 kcal/ mol. With complementarity being the only proofreading step, the estimated error frequency would be one wrong amino acid in every 100 amino acids incorporated. However, error frequencies observed in...
10.4K
Machines: Problem Solving II
310
Machines are complex structures consisting of movable, pin-connected multi-force members that work together to transmit forces. Consider a lifting tong carrying a 100 kg load. It comprises movable sections DAF and CBG linked together with member AB.
310
Language and Cognition
346
Language serves as a bridge between ideas and communication, influencing how individuals perceive and interact with the world. Psychologists have long debated whether language shapes thought or vice versa. This discussion gained grip with Edward Sapir and Benjamin Lee Whorf in the 1940s, who proposed that language determines thought, a concept known as linguistic determinism. They suggested that the vocabulary and structure of a language influence how its speakers think and perceive reality.
346
Machines: Problem Solving I
324
A toggle clamp is a mechanical device commonly used for holding and clamping objects in various applications, such as woodworking, metalworking, and assembly operations. Consider a toggle clamp subjected to a force of 200 N at the handle. The vertical clamping force can be calculated, provided the dimensions of the toggle clamp are known.
The toggle clamp system is a machine structure consisting of movable, pin-connected multi-force members that form a stabilized system to transmit forces. The...
The toggle clamp system is a machine structure consisting of movable, pin-connected multi-force members that form a stabilized system to transmit forces. The...
324
Stereotype Content Model
14.7K
The Stereotype Content Model (SCM) was first proposed by Susan Fiske and her colleagues (Fiske, Cuddy, Glick & Xu, 2002; see also Fiske, 2012 and Fiske, 2017). The SCM specifies that when someone encounters a new group, they will stereotype them based on two metrics: warmth—or that group’s perceived intent, and how likely they are to provide help or inflict harm—and competence—or their ability to carry out that objective. Depending on the warmth-competence...
14.7K
Language Development
366
Children master language quickly and with relative ease, supported by both biological predisposition and reinforcement. B. F. Skinner (1957) proposed that language is learned through reinforcement, while Noam Chomsky (1965) argued that language acquisition mechanisms are biologically determined.
The critical period for language acquisition suggests that the ability to acquire language is at its peak early in life. As people age, this proficiency decreases. Language development begins very...
The critical period for language acquisition suggests that the ability to acquire language is at its peak early in life. As people age, this proficiency decreases. Language development begins very...
366


