贝叶斯优化
第九部分:偏好是什么?
EN

社会、情感与发展心理学

第 37 章考察的是人在两个选项之间做决定的那几秒钟。本章转向心理学的其他分支,它们在更长的时间跨度、更丰富的场景中研究人:在他人的注视下、追求目标时、作为消费者、处于某种心境中、面对道德权衡、贯穿一生,以及身处房间与城市之中。人是否怀有稳定的偏好、能否通过一连串比较找到,每个分支都有自己的看法。

贯穿这些分支的有三项发现。其一,几个著名效应曾被用来论证回答彼此依赖,如今在多实验室检验中未能复现,或在校正发表偏倚后缩小到接近零。其二,人们自我报告的宽泛倾向相当稳定,而有激励的选择与单次判断是最不稳定的测量。其三,与偏好贝叶斯优化最接近的证据来自道德偏好引出:人们在不同会话中回答同一个成对问题时,有 6% 至 20% 的回答会翻转;在偏好不稳定或模型类别有误的模拟中,主动选择的查询可能并不优于随机查询。第 38.12 节汇总了这些结果所指向的观测模型;全章使用的各种标尺在第 37.1 节中有解释。

38.1 社会心理学 #

社会心理学提出的说法涉及选择、人们无法报告的态度以及他人的影响。认知失调理论(cognitive dissonance theory)认为,行为与态度冲突时人会感到不适,并通过改变态度来减轻这种不适;长期以来,它被用来解释被选中的选项为何在选择之后升值(第 37.2.2 节)。双重态度(dual attitudes)模型认为,人可以同时对同一对象持有内隐态度和外显态度(Wilson 等,2000),因此用户可能有“两个真实的偏好”(two true preferences)。此外,Asch 的从众实验通常被概括为:约 75% 的被试至少有一次附和了错误的多数(Asch,1956);在信息级联(information cascades)中,人们照搬先前的选择,而不使用自己的信息(Bikhchandani 等,1992)。这两者都暗示,向用户展示他人的评分,可能制造虚假的共识。

38.1.1 选择改变偏好,从婴儿期开始 #

除了 d=0.40d = 0.40、95% 置信区间为 [0.32,0.49][0.32, 0.49] 的无伪影元分析(第 37.2.2 节),Silver 等人(2020)在 7 项实验(N=189N = 189)中发现,尚不会说话的婴儿会回避自己没有选择的选项,并排除了新奇性与先前态度的解释:早在需要捍卫的成熟自我概念形成之前,选择就已在塑造偏好。通向失调的另一条经典路径未能复现。在诱导服从范式(induced-compliance paradigm)中,人们自愿或按指示写一篇反对自己观点的文章;该理论预测,当人们觉得写作出于自己的选择时,态度改变会更大。Vaidis 等人(2024)在 19 个国家的 39 个实验室中(N=4,898N = 4{,}898)发现,高选择条件与低选择条件之间的态度没有差异,尽管与写中性文章相比,写反态度文章确实改变了态度。这一结果支持以序贯抽样与再评价(第 37.2.2 节),而非经典的失调,作为其中的机制。

38.1.2 内隐态度与从众 #

内隐测量(例如内隐联想测验,它根据人们配对概念的快慢来推断态度)对个体的预测很差。在 217 份报告(N=36,071N = 36{,}071)中,内隐测量与外显测量对预测群际行为各有独特贡献,分别为 β=.14\beta = .14 与 .11.11,相关高度异质(Kurdi 等,2019)。一项涵盖 492 项研究的网络元分析发现,内隐测量可以改变,但改变幅度很小(∣d∣<.30|d| < .30),而且内隐的改变并不中介外显或行为的改变(Forscher 等,2019)。内隐测量的时间稳定性(加权平均 r=.54r = .54)低于外显测量(r=.75r = .75)(Gawronski 等,2017)。从众得到了复现:在一项有五名实验同谋的 Asch 复现中(N=210N = 210),标准线段判断任务的错误率为 33%,金钱激励将其降至 25%(Franzen 与 Mader,2023)。

38.1.3 作为信念更新的社会影响 #

如今,社会影响被理解为按不确定性加权的信念更新。在一个 14 至 24 岁人群的大样本中,在延迟折扣(delay discounting)任务中(在较早的较小奖励与较晚的较大奖励之间做选择),采纳他人偏好的倾向源于对自身偏好更大的不确定性;这种不确定性以及随之而来的易受影响程度,在 1.5 年中有所下降(Reiter 等,2021)。人工智能如今也成了影响的来源之一:人们与一个有偏差的人工智能系统交互时,把面孔阵列判断为悲伤多于快乐的比例,从基线时的 49.9% 上升到 56.3%(d=0.84d = 0.84)(Glickman 与 Sharot,2025)。

38.1.4 对观测模型的含义 #

  • 保留选择历史项,并配以再评价机制。这一项可以根据选择的难度与反应时预测漂移的大小,两者都可由界面记录(推断)。
  • 不要过早展示他人的选择。他人的选择、受欢迎程度或系统推荐的默认选项,不应出现在会话早期,因为此时用户自身的不确定性最高,易受影响的程度也随之最高(推断)。第 20.5 节讨论了群体模型,这类模型汇集众人的信息,而不把一个人的回答展示给另一个人。
  • 偏好的不确定性可以测量。重复配对上的不一致可以估计这种不确定性,它既可作为噪声尺度,也可作为此人易受影响的警示(推断)。
  • 内隐测量不是第二条偏好通道。证据更充分的辅助信号是反应时(推断),见第 39.3 节的报告。
第 38.1 节引用的文献 11
  1. Wilson 等人(2000)A model of dual attitudes
  2. Asch(1956)Studies of independence and conformity: I. A minority of one against a unanimous majority
  3. Bikhchandani 等人(1992)A Theory of Fads, Fashion, Custom, and Cultural Change as Informational Cascades
  4. Silver 等人(2020)When Not Choosing Leads to Not Liking: Choice-Induced Preference in Infancy
  5. Vaidis 等人(2024)A Multilab Replication of the Induced-Compliance Paradigm of Cognitive Dissonance
  6. Kurdi 等人(2019)Relationship between the Implicit Association Test and intergroup behavior: A meta-analysis
  7. Forscher 等人(2019)A meta-analysis of procedures to change implicit measures
  8. Gawronski 等人(2017)Temporal Stability of Implicit and Explicit Measures: A Longitudinal Analysis
  9. Franzen 与 Mader(2023)The power of social influence: A replication and extension of the Asch experiment
  10. Reiter 等人(2021)Preference uncertainty accounts for developmental effects on susceptibility to peer influence in adolescence
  11. Glickman 与 Sharot(2025)How human–AI feedback loops alter human perceptual, emotional and social judgements

38.2 动机与目标 #

关于偏好的层级观认为,终极偏好是稳定的,工具性偏好则在情境中建构。在控制理论中,行为是一个目标层级,越高层的目标越抽象、变化越慢(Carver 与 Scheier,1982);在自我决定理论(self-determination theory)中,驱动动机的是对自主、胜任与关系的需要(Deci 与 Ryan,2000);终极价值则据称是稳定的(Schwartz,1992)。如果终极目标稳定,偏好贝叶斯优化就有稳定的对象可以寻找。

38.2.1 习惯、需要与价值观 #

习惯与目标的关系仍无定论:Wood 等人(2022)认为习惯独立于目标运作,然而 de Wit 等人(2018)报告,在以人为对象的实验室研究中,五次通过过度训练诱导习惯的尝试都失败了。对大多数人来说,价值观层级是稳定的:一项针对早期青少年、为期两年的研究中,75% 的人两个时间点的价值观层级相关至少为 .85,5% 的人至多为 .12(Vecchione 等,2020)。但价值是相对于目标而言的。当被试必须选出最差而非最好的选项时,被试的行为以及通常被解读为价值的神经活动,主要受目标一致性(goal congruency)支配,即选项在多大程度上服务于当前目标,而非受期望奖励支配(Frömer 等,2019)。

38.2.2 对观测模型的含义 #

  • 效用相对于任务框架而言。挑选自己喜欢的、挑选适合客户的、挑选要排除的,是不同的目标;因此指导语应当保持不变并记录在案,未经检验,“哪个更差?”不应与“哪个更好?”共用一个似然(推断)。
  • 反应习惯是干扰,而非偏好。总是选左边的选项、总是保留当前最优点,都应归入针对个人的干扰项(推断)。
  • 自主为限制影响提供了理由,第 41.2 节将展开讨论这一关切(推断)。
第 38.2 节引用的文献 7
  1. Carver 与 Scheier(1982)Control theory: A useful conceptual framework for personality–social, clinical, and health psychology
  2. Deci 与 Ryan(2000)The "What" and "Why" of Goal Pursuits: Human Needs and the Self-Determination of Behavior
  3. Schwartz(1992)Universals in the Content and Structure of Values: Theoretical Advances and Empirical Tests in 20 Countries
  4. Wood 等人(2022)Habits and Goals in Human Behavior: Separate but Interacting Systems
  5. de Wit 等人(2018)Shifting the balance between goals and habits: Five failures in experimental habit induction
  6. Vecchione 等人(2020)Stability and change of basic personal values in early adolescence: A 2‐year longitudinal study
  7. Frömer 等人(2019)Goal congruency dominates reward value in accounting for behavioral and neural correlates of value-based decision-making

38.3 消费者心理学 #

消费者研究认为,做选择以及反复选择,会使偏好更稳定(Hoeffler 与 Ariely,1999),早期经历能有力地预测最终偏好(Hoeffler 等,2006);照此推论,一次会话最初的候选将决定哪些偏好结晶下来。选择过载(choice overload)是指选项过多会损害选择:在一项著名的现场研究中,陈列 24 种果酱时有 3% 的顾客购买,陈列 6 种时则有 30%(Iyengar 与 Lepper,2000);但一项涵盖 63 个条件的元分析发现,平均效应接近零,方差很大(Scheibehenne 等,2010)。

38.3.1 选择过载与偏好结晶 #

选择过载确实存在,但高度依赖条件:一项多层多变量元分析表明,这一效应在六个因变量和四个调节变量之间差异很大,且这些调节变量相互作用(McShane 与 Böckenholt,2018)。人们从 6、12 或 24 个物品中做选择时,纹状体与前扣带皮层的活动呈倒 U 形,峰值在 12,被试也认为这一数量大致合适;而只浏览、不选择时,这一模式便消失了(Reutskaja 等,2018)。我们没有找到偏好结晶的直接复现。在一项结合眼动追踪的重复离散选择实验中,不稳定回答者与稳定回答者的潜在偏好并无差异,不稳定反映的是效用接近时做选择的难度(Fraser 等,2021)。

38.3.2 对观测模型的含义 #

  • 不一致是噪声尺度,而非用户类型。接近无差异时、面对复杂选项时,不一致应当上升(推断)。
  • 画廊也许有最佳大小,但没有通用的最佳大小。多选项界面(第 20.3 节)的选项数可能存在内部最优值,它取决于用户是必须做选择还是可以只浏览;12 只是某一个领域中的结果(推断)。
  • 随机化并记录初始设计。早期经历与选择都会影响会话的终点;在不同用户之间随机化起始设计,就能把优化器的作用与用户的偏好区分开(推断)。照片增强案例研究(第 25 章)中的起始图像设置就可以这样随机化。
第 38.3 节引用的文献 7
  1. Hoeffler 与 Ariely(1999)Constructing Stable Preferences: A Look Into Dimensions of Experience and Their Impact on Preference Stability
  2. Hoeffler 等人(2006)Path dependent preferences: The role of early experience and biased search in preference development
  3. Iyengar 与 Lepper(2000)When choice is demotivating: Can one desire too much of a good thing?
  4. Scheibehenne 等人(2010)Can There Ever Be Too Many Options? A Meta-Analytic Review of Choice Overload
  5. McShane 与 Böckenholt(2018)Multilevel Multivariate Meta-analysis with Application to Choice Overload
  6. Reutskaja 等人(2018)Choice overload reduces neural signatures of choice set value in dorsal striatum and anterior cingulate cortex
  7. Fraser 等人(2021)Preference stability in discrete choice experiments. Some evidence using eye-tracking

38.4 情绪与享乐心理学 #

情绪研究提供了三种说法。一是享乐适应(hedonic adaptation):一项著名的研究发现,彩票中奖者并不比其他人更快乐(Brickman 等,1978),因此早期的偏好配对可能会过期。二是情感即信息(feelings as information):人们在判断时会参考自己当前的心境,例如生活满意度评分会受天气左右(Schwarz 与 Clore,1983),因此会话可能需要一段“情绪冷却”。三是双过程理论,它认为快速的强制选择捕捉的是直觉偏好。

38.4.1 三处修正 #

这三种说法都需要修正。一项预注册分析考察了中奖后 5 至 22 年的瑞典彩票玩家,发现生活满意度上升,并在十多年里保持在较高水平,没有消退的迹象,而对快乐与心理健康的影响则显著较小(Lindqvist 等,2020);在 18 种生活事件中,情感在两年内适应了每一种积极事件,但评价性幸福感从经济收益和退休中获得了持久的好处(Kettlewell 等,2020)。在九项样本更大(NN 从 118 至 401)的直接复现与概念复现中,心境对生活满意度判断的影响大多不显著,即使显著,也比早先的结果小得多(Yap 等,2017)。汇总 21 项预注册复现的结果,比较必须快速决定的人与被迫等待的人对共同项目的贡献:两组之差在意向治疗分析(按分配的组别进行比较)中为 −0.37-0.37 个百分点,而原始数据中为 8.6 个百分点;时间压力组中有 65.9% 的人未能在时限内作出决定(Bouwmeester 等,2017)。

38.4.2 对观测模型的含义 #

  • 偶发心境只是微弱的噪声来源。情绪冷却和按心境安排时机的探索都缺乏支持(推断)。
  • 两种力量以相反方向作用于反复出现的选项。选择引起的再评价会抬高不断获胜的当前最优点,而重复消费带来的享受下降(Galak 与 Redden,2018)会压低它;漂移项的符号应当估计,而不是固定(推断)。
  • 快速的强制选择并不能揭示直觉的自我(推断)。
  • 会话结束时的满意度测量的是早期情感。当下的喜爱与是否合用会随时间分化,因此需要延迟随访(推断),这也是第 46.7 节的建议。
第 38.4 节引用的文献 7
  1. Brickman 等人(1978)Lottery winners and accident victims: Is happiness relative?
  2. Schwarz 与 Clore(1983)Mood, misattribution, and judgments of well-being: Informative and directive functions of affective states
  3. Lindqvist 等人(2020)Long-Run Effects of Lottery Wealth on Psychological Well-Being
  4. Kettlewell 等人(2020)The differential impact of major life events on cognitive and affective wellbeing
  5. Yap 等人(2017)The effect of mood on judgments of subjective well-being: Nine tests of the judgment model
  6. Bouwmeester 等人(2017)Registered Replication Report: Rand, Greene, and Nowak (2012)
  7. Galak 与 Redden(2018)The Properties and Antecedents of Hedonic Decline

38.5 道德心理学 #

道德心理学既提出了关于偏好结构的说法,也提供了关于重复成对引出的最直接证据。神圣价值(sacred values)抗拒权衡:用金钱交换神圣价值的提议会激起愤怒(Tetlock 等,2000),由此产生的论证是,在偏好贝叶斯优化中,神圣价值应当作为硬约束。第二种论证认为,相继的道德回答会相互补偿。在道德许可(moral licensing)中,一件善行会为之后的失范开出许可(Monin 与 Miller,2001);在自我概念维持(self-concept maintenance)中,诚实的人只在仍能自视为诚实之人的限度内作弊,而关于道德的提醒会减少作弊(Mazar 等,2008)。

38.5.1 诚实、许可与撤稿 #

自我概念维持有 25 项直接复现(N=5,786N = 5{,}786),其主要元分析(19 项复现,n=4,674n = 4{,}674)得到 d=−0.04d = -0.04,而原始结果为 d=0.48d = 0.48(Verschuere 等,2018)。Journal of Marketing Research 于 2024 年对那篇 2008 年的论文发布了关注声明(Journal of Marketing Research,2024);同一组作者 2012 年发表的一篇论文,内容是在表格顶部签署诚信承诺,已被撤稿(Shu 等,2012)。

补偿论证需要道德许可以个体内的形式成立,但它并不以这种形式成立。Rotella 等人(2026)汇总了 115 项实验(N=21,770N = 21{,}770)。未经校正的多层估计为 Hedges' g=0.21g = 0.21(gg 是经过小样本校正的 Cohen's dd),但经过偏倚校正的稳健贝叶斯元分析给出的估计接近零(gg 在 −0.08-0.08 与 −0.02-0.02 之间),并有非常强的发表偏倚证据。这一效应取决于是否有人注视:被试明确处于被观察状态时 g=0.65g = 0.65,否则 g=0.13g = 0.13;作者把这解读为一种人际的、关乎声誉的效应。对最初的道德凭证研究所做的注册复现(N=932N = 932)不支持存在一致的效应(Xiao 等,2024)。图 38.1 将这些结果与经受住检验的选择引起的改变并列。

原始估计或未校正估计后续检验或校正后估计报告的区间小中大道德提醒与作弊,25 项复现道德许可,115 项实验被观察时的道德许可未被观察时的道德许可早晨道德效应(预印本)选择引起的改变,43 项研究−0.20.00.20.40.60.8效应量 d(或 g)第 1 组第 2 组重叠 81%d = 0.48:两个分布重叠 81%;均值较高的一组在 63% 的情况下胜过较低的一组。
原始估计或未校正估计后续检验或校正后估计报告的区间小中大道德提醒与作弊,25 项复现道德许可,115 项实验被观察时的道德许可未被观察时的道德许可早晨道德效应(预印本)选择引起的改变,43 项研究−0.20.00.20.40.60.8效应量 d(或 g)第 1 组第 2 组重叠 81%d = 0.48:两个分布重叠 81%;均值较高的一组在 63% 的情况下胜过较低的一组。
图 38.1 曾用于论证回答彼此依赖的效应,置于同一标尺上。空心点为原始估计或未校正估计,实心点为后续估计、经偏倚校正的估计或子组估计,横条为报告的区间或范围。点击某一行或移动滑块,下方面板会画出两个均值相差 d 的正态分布;这是正态模型的性质,而非数据。

可以尝试以下几点:

  • 原始的道德提醒效应 d=0.48d = 0.48 时,两个分布重叠 81%,较高一组中的人在 63% 的情况下得分高于较低一组中的人。复现得到的 −0.04-0.04 时,两者重叠 98%。
  • 被观察时的道德许可,g=0.65g = 0.65,重叠 75%;未被观察时,g=0.13g = 0.13,重叠 95%。

38.5.2 神圣价值与道德困境 #

神圣价值既不普遍,也不固定。看到某种神圣价值被用于牟利,会降低它的神圣性以及人们对权衡的抗拒(7 项研究,N=2,785N = 2{,}785)(Ruttan 与 Nordgren,2021)。在离散选择实验中,把对禁忌权衡的惩罚加入随机效用模型后,潜在类别分析显示,有些群体视这些权衡为禁忌,另一些则不然;忽略这种厌恶,会把为拯救生命的支付意愿高估约 3.5 倍(Smeele 等,2025)。此外,假设性的道德判断不能预测真实行为:人们对电车式困境的回答,无法预测他们是否真的会电击一只老鼠,以使另外五只免受电击(Bostyn 等,2018)。

38.5.3 道德回答有多稳定?#

汇总后的功利主义判断在不同时间、八种情境之间都高度一致(Helzer 等,2017)。单次判断则不然:在相隔 6 至 8 天的两轮调查之间,对某一牺牲困境,平均有 49% 的被试改变了评分,而这些改变中只有约 17% 来自自称改变了想法的人(Rehren 与 Sinnott-Armstrong,2022)。有两项研究反复提出同一个成对问题。一项研究在两周内十次询问被试:两位患者中谁应当得到唯一可用的肾脏。对于有争议的情形,人们约有 10% 至 18% 的回答发生了改变,选择越慢、越难,改变越频繁(Boerstler 等,2024)。在另一项规模更大的研究中,400 多名被试在 3 至 5 次会话中回答成对的肾脏分配比较;他们对同一情形的回答平均有 6% 至 20% 发生了改变,简单模型的预测表现随不稳定程度增加、随时间推移而下降(Keswani 等,2026)。可以在图 38.2 中做一个简短版本的测试。

第 1 轮(共 2 轮)· 第 1 题(共 8 题)现有一个肾脏可供移植。应该给哪一位患者?患者 1年龄30 岁需抚养的子女0可获得的生命年约 12 年患者 2年龄55 岁需抚养的子女2可获得的生命年约 9 年患者是假设的。两人在医学上都匹配。
第 1 轮(共 2 轮)· 第 1 题(共 8 题)现有一个肾脏可供移植。应该给哪一位患者?患者 1年龄30 岁需抚养的子女0可获得的生命年约 12 年患者 2年龄55 岁需抚养的子女2可获得的生命年约 9 年患者是假设的。两人在医学上都匹配。
图 38.2 同一个分配问题,问两次。回答八个问题:两位假设的患者中,谁应当得到唯一可用的肾脏;第 2 轮按新的顺序再问一遍,并交换左右位置。总结显示有多少回答发生了改变,并与重复间隔数天时测得的比例并列。你的回答只保存在本页面;八个问题仅相隔几分钟,只是演示,而非测量。

与偏好贝叶斯优化最相关的结果涉及主动学习,即像采集函数那样,让每个后续问题尽可能提供信息(第 19.3 节)。Keswani 等人(2024)指出了主动学习的三个前提:偏好稳定且不受提问顺序影响,假设类别正确,噪声有限。在违反这些前提的模拟中,主动学习在某些设定下与随机提问表现相当或更差;只有当不稳定性和噪声都很小、偏好可由假设类别近似表示时,它才仍然值得使用。

38.5.4 噪声还是改变?#

6% 至 20% 的翻转率,可能是固定偏好在回答时带有噪声,也可能是偏好本身在会话之间移动,用 Keswani 等人(2026)的话说,即可能的“道德改变”(moral change)。噪声应当通过平均消除,改变则应当追踪。设两个选项之间的长期效用差为 Δ\Delta,每次会话中偏移 δ∼N(0,τ2)\delta \sim \N(0, \tau^2),每个选项带有尺度为 σ\sigma 的概率单位反应噪声(第 16.3 节)。同一会话中的两个回答共享 δ\delta,两次会话中的回答则不共享。记 pδ=Φ((Δ+δ)/(2 σ))p_\delta = \Phi\big((\Delta + \delta)/(\sqrt{2}\,\sigma)\big),两个回答不同的概率为

Pwithin(Δ)=Eδ ⁣[2 pδ (1−pδ)],Pacross(Δ)=2 pˉ (1−pˉ),pˉ=Φ ⁣(Δ2σ2+τ2),P_{\text{within}}(\Delta) = \E_\delta\!\left[2\,p_\delta\,(1 - p_\delta)\right], \qquad P_{\text{across}}(\Delta) = 2\,\bar p\,(1 - \bar p),\quad \bar p = \Phi\!\left(\frac{\Delta}{\sqrt{2\sigma^2 + \tau^2}}\right),
(38.1)

因为两次独立会话各自以平均概率 pˉ\bar p 回答“是”。当 τ>0\tau > 0 时,跨会话的回答比同一会话内的回答更常不同,如图 38.3 所示。

在一次会话中问两次在两次会话中各问一次0%10%20%30%40%50%P(两次回答不同)0.00.51.01.52.02.53.0两个选项之间的长期效用差距 Δ跨会话的报告值6% 至 20%16%16%对各对取平均差距 0 至 1.5
在一次会话中问两次在两次会话中各问一次P(两次回答不同)0%10%20%30%40%50%0123效用差距 Δ报告值6% 至 20%16%16%对各对取平均差距 0 至 1.5
图 38.3 噪声还是改变?对同一对的两次回答不同的概率随长期效用差距 Δ 的变化:在一次会话中问两次(蓝色),或在两次会话中各问一次(橙色,虚线);σ 是每个选项的反应噪声,τ 是会话之间的改变。侧面板在阴影范围内对各对取平均,并与肾脏分配研究报告的 6% 至 20% 并列。会话内的重复忽略了记忆的作用;这一模型仅作示意。

τ=0\tau = 0、σ=0.3\sigma = 0.3 时,两条曲线重合,平均翻转率约为 16%,落在报告的范围之内:单纯的噪声已足以解释。σ=0.05\sigma = 0.05、τ=0.4\tau = 0.4 时,跨会话的平均值同样约为 15%,但在一次会话之内降到约 3%。只有把会话内的重测与之后某次会话中的重测相比较,才能区分这两种情况;这正是第 37.4.4 节建议要么立即重测、要么很久之后再重测的原因。

38.5.5 对观测模型的含义 #

  • 以随机查询为对照检验采集函数。在分配、政策与安全权衡这类涉及价值的领域中,应在留出的重复配对上把采集函数与随机查询相比较,并在每次会话中保留固定比例的随机查询或重复查询(推断);第 28.9 节中的比较很少包含这样的对照。
  • 预期一致性存在下限。跨会话 6% 至 20% 的翻转率需要引入会话层面的随机效应;一次会话之内的收敛,不应外推到多次会话(推断)。
  • 用针对个人的惩罚取代硬约束:对受保护属性的权衡施加惩罚,并将其置于用户类别的混合模型中(推断)。
  • 放弃许可效应,并在引用证据之前先加核查。道德许可只在被观察时出现,因此基于道德许可的补偿不适用于私密会话;谁能看到回答,由此成为一个实验因素。以某个行为效应为设计选择辩护的论文,应当核查该效应的复现与撤稿状况(推断)。
  • 假设性的验证不是真正的验证(推断)。道德领域中的偏好贝叶斯优化研究是第 38.13 节列出的缺口之一;这类研究应当以真实决策来验证,而不只以更多的假设性回答来验证。
第 38.5 节引用的文献 16
  1. Tetlock 等人(2000)The psychology of the unthinkable: Taboo trade-offs, forbidden base rates, and heretical counterfactuals
  2. Monin 与 Miller(2001)Moral credentials and the expression of prejudice
  3. Mazar 等人(2008)The Dishonesty of Honest People: A Theory of Self-Concept Maintenance
  4. Verschuere 等人(2018)Registered Replication Report on Mazar, Amir, and Ariely (2008)
  5. Journal of Marketing Research(2024)Expression of Concern:“The Dishonesty of Honest People: A Theory of Self-Concept Maintenance”
  6. Shu 等人(2012)RETRACTED: Signing at the beginning makes ethics salient and decreases dishonest self-reports in comparison to signing at the end
  7. Rotella 等人(2026)Observation Moderates the Moral Licensing Effect: A Meta-Analytic Test of Interpersonal and Intrapsychic Mechanisms
  8. Xiao 等人(2024)Licensing via Credentials: Replication Registered Report of Monin and Miller (2001) with Extensions Investigating the Domain-Specificity of Moral Credentials and the Association Between the Credential Effect and Trait Reputational Concern
  9. Ruttan 与 Nordgren(2021)Instrumental use erodes sacred values
  10. Smeele 等人(2025)Taboo trade-off aversion in choice behaviors: A discrete choice model and application to health-related decisions
  11. Bostyn 等人(2018)Of Mice, Men, and Trolleys: Hypothetical Judgment Versus Real-Life Behavior in Trolley-Style Moral Dilemmas
  12. Helzer 等人(2017)Once a Utilitarian, Consistently a Utilitarian? Examining Principledness in Moral Judgment via the Robustness of Individual Differences
  13. Rehren 与 Sinnott-Armstrong(2022)How Stable are Moral Judgments?
  14. Boerstler 等人(2024)On The Stability of Moral Preferences: A Problem with Computational Elicitation Methods
  15. Keswani 等人(2026)Moral Change or Noise? On Problems of Aligning AI With Temporally Unstable Human Feedback
  16. Keswani 等人(2024)On the Pros and Cons of Active Learning for Moral Preference Elicitation

38.6 人格与毕生发展 #

一项被广泛引用的元分析发现,人格特质的等级顺序稳定性从儿童期的 .31 上升到 30 岁时的 .64,并在 50 至 70 岁之间达到约 .74 的平台(Roberts 与 DelVecchio,2000);由此有人建议,将中年用户的数据视为更可靠。

38.6.1 什么是稳定的,从何时起 #

一项元分析汇总了 2005 年以来发表的纵向研究,涵盖等级顺序稳定性(189 项研究,N=178,503N = 178{,}503)与均值水平变化(276 项研究,N=242,542N = 242{,}542),发现等级顺序稳定性在成年早期即达到平台,几乎没有证据表明 25 岁之后还会进一步上升(Bleidorn 等,2022)。陈述的、类似特质的偏好相当稳定:1,507 名成年人完成了 39 种风险偏好测量,陈述测量与行为测量之间只有微弱的相关,但陈述测量构成一个一般因子,在六个月内高度可靠(Frey 等,2017)。全球偏好调查(Global Preferences Survey)在 76 个国家调查了 80,000 人,发现国家内部的异质性大于国家之间的异质性(Falk 等,2018)。

38.6.2 对观测模型的含义 #

  • 群体先验有帮助,但每个用户仍然需要单独学习,因为国家内部的异质性很大(推断);第 32.5 节报告了群体先验在交互式系统中的表现。
  • 在比较之外辅以几个陈述题项。成对选择属于不太可靠的行为类测量(第 37.1 节)。陈述题项可以承载稳定成分,例如令高斯过程的先验均值依赖于这些题项(推断;这种实现来自高斯过程框架)。
  • 报告跨会话的重测一致性,因为一次会话之内的收敛只是稳定偏好的微弱证据(推断)。
第 38.6 节引用的文献 4
  1. Roberts 与 DelVecchio(2000)The rank-order consistency of personality traits from childhood to old age: A quantitative review of longitudinal studies
  2. Bleidorn 等人(2022)Personality stability and change: A meta-analysis of longitudinal studies
  3. Frey 等人(2017)Risk preference shares the psychometric structure of major psychological traits
  4. Falk 等人(2018)Global Evidence on Economic Preferences*

38.7 进化心理学 #

择偶偏好中的性别差异最初是在 37 种文化中报告的(Buss,1989),后来在 45 个国家(N=14,399N = 14{,}399)中得到复现(Walter 等,2020),但这些差异对个体几乎说明不了什么。一项注册报告(N=10,358N = 10{,}358,43 个国家)发现,陈述的理想伴侣偏好在总体模式上与对伴侣的评价相符(β=.19\beta = .19),但针对具体特质的效应平均只有 β=.04\beta = .04(Eastwick 等,2025)。因此,若先验依据用户对各属性所陈述的重要性来初始化,就应当打折扣,或用显示出的比较加以校正(推断)。

第 38.7 节引用的文献 3
  1. Buss(1989)Sex differences in human mate preferences: Evolutionary hypotheses tested in 37 cultures
  2. Walter 等人(2020)Sex Differences in Mate Preferences Across 45 Countries: A Large-Scale Replication
  3. Eastwick 等人(2025)A worldwide test of the predictive validity of ideal partner preference matching

38.8 发展心理学与行为遗传学 #

遗传力(heritability)指人群中某一特质的变异里,与遗传差异相关联的部分所占的比例。在 9,169 名瑞典双胞胎中,遗传效应解释了音乐奖赏敏感性至多 54% 的方差(Bignardi 等,2025);而个人对面孔的偏好主要来自每个人独特的环境(Germine 等,2015)。遗传力本身并不决定先验的强度;模型需要的是一种划分:哪部分是人与人之间共享的成分,哪部分是个人特有的成分,而共享部分的权重本身也可以是个人层面的参数(推断)。选择引起的改变从婴儿期就已存在(第 38.1.1 节),因此招募更善于反思的用户并不能避免它(推断)。

第 38.8 节引用的文献 2
  1. Bignardi 等人(2025)Twin modelling reveals partly distinct genetic pathways to music enjoyment
  2. Germine 等人(2015)Individual Aesthetic Preferences for Faces Are Shaped Mostly by Environments, Not Genes

38.9 睡眠与昼夜节律 #

早晨道德效应(morning morality effect)指人在早晨更诚实(Kouchaki 与 Smith,2014),有人援引它,建议在一个人的昼夜节律高峰时引出偏好。一项 N=1,006N = 1{,}006 的概念复现得到的比值比为 1.04(95% 置信区间为 [0.93,1.17][0.93, 1.17]),一项元分析得到 d=0.04d = 0.04(Zickfeld 等,2024),这是一篇 2024 年的预印本。自我报告的睡眠更适合作为噪声尺度与失误率的协变量,而不是作为效用上带符号的偏差(推断)。

第 38.9 节引用的文献 2
  1. Kouchaki 与 Smith(2014)The Morning Morality Effect
  2. Zickfeld 等人(2024)Investigating the Morning Morality Effect and its Mediating and Moderating Factors

38.10 神经多样性 #

在一项预注册的对抗式合作中,自述有孤独症诊断的人对与结果无关的特征学得少得多;在整个样本中,这种减少随孤独症特质的程度而变化,呈现维度式而非类别式的模式(Ben-Artzi 等,2026)。因此,位置、顺序、无关视觉特征之类的干扰偏差,需要从交互本身估计个人层面的干扰参数;这样做可以适应每个用户,而无需收集任何诊断信息(推断)。

第 38.10 节引用的文献 1
  1. Ben-Artzi 等人(2026)Autism-associated learning patterns show reduced credit assignment to outcome-irrelevant features

38.11 环境心理学 #

瞭望庇护理论(prospect-refuge theory)认为,人们偏好既能提供视野、又能提供庇护的地方,但支持它的定量证据并不一致(Dosen 与 Ostwald,2016)。798 名被试为 200 个室内空间评分,连贯性(coherence)、吸引力(fascination)与居家感(hominess)三个成分解释了 90% 的方差,这一结构还在一个独立样本(n=614n = 614)中得到复现(Coburn 等,2020);这三个成分可以在多目标表述中充当可解释的目标(第 14.5 节)(推断)。81,630 名志愿者对来自 56 个城市的街景图像做了 117 万次成对比较(Dubey 等,2016);如果限定于某种文化或某个城市,这些数据可以为城市设计提供群体先验(推断)。第 44.2 节与第 44.3 节讨论建筑与城市的设计。

第 38.11 节引用的文献 3
  1. Dosen 与 Ostwald(2016)Evidence for prospect-refuge theory: a meta-analysis of the findings of environmental preference research
  2. Coburn 等人(2020)Psychological and neural responses to architectural interiors
  3. Dubey 等人(2016)Deep Learning the City: Quantifying Urban Perception at a Global Scale

38.12 共同的模式 #

同一模式在各个分支中反复出现(表 38.1):宽泛的自我报告是最稳定的偏好测量,单次选择与有激励的行为最不稳定,而群体规律对个人偏好的具体形状几乎没有预测力。

表 38.1 按测量类型划分的偏好稳定性,涵盖本章与上一章的各个分支。
领域 宽泛的自我报告 单次选择、行为或具体匹配
风险 陈述的倾向:随时间的平均信度 0.61 彩票选择等行为测量:0.25
道德 汇总的功利主义判断随时间保持一致 49% 的人在几天内改变对一个困境的评分;6% 至 20% 的成对回答在会话之间翻转
伴侣 陈述的理想在总体模式上与评价相符,β=.19\beta = .19 针对具体特质的匹配 β=.04\beta = .04
态度 外显测量:稳定性 r=.75r = .75 内隐测量:稳定性 r=.54r = .54

结合第 37 章,这些结果提示了如下形式的观测模型(推断)。设一个人在会话 ss 中的效用为

us(x)=m(x)+f(x)+ds(x),u_s(\vx) = m(\vx) + f(\vx) + d_s(\vx),
(38.2)

其中 mm 是群体均值,由群体数据或少数几项陈述偏好提供信息;ff 是此人的稳定成分,即第 18 章的高斯过程;dsd_s 是方差较小的会话层面偏差,使效用可以在会话之间移动,而在一次会话之内保持不变。于是,第 tt 步的比较服从

P(x≻x′)=λlapse2+(1−λlapse) Φ ⁣(us(x)−us(x′)+b ct2 σ),\Prob(\vx \succ \vx') = \frac{\lambda_{\text{lapse}}}{2} + (1 - \lambda_{\text{lapse}})\, \Phi\!\left(\frac{u_s(\vx) - u_s(\vx') + b\, c_t}{\sqrt{2}\,\sigma}\right),
(38.3)

其中 λlapse\lambda_{\text{lapse}} 是失误率(给出与选项无关的回答的概率),ct=±1c_t = \pm 1 编码哪个选项在左边,bb 是针对个人的位置偏差,σ\sigma 是噪声尺度,可能依赖于此人的不确定性以及选项之间的接近程度。重复配对为 σ\sigma 与 λlapse\lambda_{\text{lapse}} 提供信息,平衡过的位置为 bb 提供信息,之后会话中的重复为 dsd_s 的方差提供信息。在我们找到的偏好优化方法中,没有一种实现了这一模型;它的各个部分是分别检验过的,大多在优化之外(推断)。另有两种做法使它完整:保留一部分随机查询作为采集函数的对照,以及对结果做延迟评估,并与产生它的会话分开(推断)。

38.13 已定、有争议与缺失 #

研究现状已定、有争议与缺失

已定。选择改变偏好,从婴儿期就已开始(Silver 等,2020)。通向失调的诱导服从路径在 39 个实验室中未能复现(Vaidis 等,2024)。道德提醒并不减少作弊(在对 25 项直接复现的主要分析中,d=−0.04d = -0.04)(Verschuere 等,2018),道德许可只在被观察时出现(Rotella 等,2026)。简单判断上的从众比例仍约为三分之一(Franzen 与 Mader,2023)。宽泛的陈述倾向稳定,单次选择则不然;人格的稳定性在约 25 岁时达到平台(Bleidorn 等,2022)。同一个道德成对问题,在会话之间有 6% 至 20% 的时候会翻转(Keswani 等,2026)。心境对幸福感判断的影响比曾经报告的小得多(Yap 等,2017),时间压力也不能揭示直觉偏好(Bouwmeester 等,2017)。

有争议。习惯是否独立于目标运作。选择过载在特定的调节变量组合之外是否存在。会话之间的翻转是噪声还是偏好的真实改变,仅凭翻转率无法判定(图 38.3)。陈述偏好与行为偏好测量的是否为同一构念。

缺失。偏好结晶的直接复现。在同一设计任务中对终极目标稳定性与工具性建构的检验。偏好贝叶斯优化在道德领域中的任何应用。这样的偏好贝叶斯优化研究:在留出的重复配对上把采集函数与随机查询相比较,对会话层面偏差建模,或估计针对个人的干扰参数。在会话之外对偏好贝叶斯优化所选设计的延迟评估。

第 38.13 节引用的文献 9
  1. Silver 等人(2020)When Not Choosing Leads to Not Liking: Choice-Induced Preference in Infancy
  2. Vaidis 等人(2024)A Multilab Replication of the Induced-Compliance Paradigm of Cognitive Dissonance
  3. Verschuere 等人(2018)Registered Replication Report on Mazar, Amir, and Ariely (2008)
  4. Rotella 等人(2026)Observation Moderates the Moral Licensing Effect: A Meta-Analytic Test of Interpersonal and Intrapsychic Mechanisms
  5. Franzen 与 Mader(2023)The power of social influence: A replication and extension of the Asch experiment
  6. Bleidorn 等人(2022)Personality stability and change: A meta-analysis of longitudinal studies
  7. Keswani 等人(2026)Moral Change or Noise? On Problems of Aligning AI With Temporally Unstable Human Feedback
  8. Yap 等人(2017)The effect of mood on judgments of subjective well-being: Nine tests of the judgment model
  9. Bouwmeester 等人(2017)Registered Replication Report: Rand, Greene, and Nowak (2012)

延伸阅读 #

参考文献

  1. Asch, S. E. (1956). Studies of independence and conformity: I. A minority of one against a unanimous majority. Psychological Monographs: General and Applied. 引用于 §38.1
  2. Ben-Artzi, I., Rozenkrantz, L., and Shahar, N. (2026). Autism-associated learning patterns show reduced credit assignment to outcome-irrelevant features. Translational Psychiatry. 引用于 §38.10
  3. Bignardi, G., Wesseldijk, L. W., Mas-Herrero, E., Zatorre, R. J., Ullén, F., Fisher, S. E., and Mosing, M. A. (2025). Twin modelling reveals partly distinct genetic pathways to music enjoyment. Nature Communications. 引用于 §38.8
  4. Bikhchandani, S., Hirshleifer, D., and Welch, I. (1992). A Theory of Fads, Fashion, Custom, and Cultural Change as Informational Cascades. Journal of Political Economy. 引用于 §38.1
  5. Bleidorn, W., Schwaba, T., Zheng, A., Hopwood, C. J., Sosa, S. S., Roberts, B. W., and Briley, D. A. (2022). Personality stability and change: A meta-analysis of longitudinal studies. Psychological Bulletin. 引用于 §38.6 §38.13
  6. Boerstler, K., Keswani, V., Chan, L., Borg, J. S., Conitzer, V., Heidari, H., and Sinnott-Armstrong, W. (2024). On The Stability of Moral Preferences: A Problem with Computational Elicitation Methods. AIES. 引用于 §38.5
  7. Bostyn, D. H., Sevenhant, S., and Roets, A. (2018). Of Mice, Men, and Trolleys: Hypothetical Judgment Versus Real-Life Behavior in Trolley-Style Moral Dilemmas. Psychological Science. 引用于 §38.5
  8. Bouwmeester, S., Verkoeijen, P. P. J. L., Aczel, B., Barbosa, F., Bègue, L., Brañas-Garza, P., … Wollbrant, C. E. (2017). Registered Replication Report: Rand, Greene, and Nowak (2012). Perspectives on Psychological Science. 引用于 §38.4 §38.13
  9. Brickman, P., Coates, D., and Janoff-Bulman, R. (1978). Lottery winners and accident victims: Is happiness relative? Journal of Personality and Social Psychology. 引用于 §38.4
  10. Buss, D. M. (1989). Sex differences in human mate preferences: Evolutionary hypotheses tested in 37 cultures. Behavioral and Brain Sciences. 引用于 §38.7
  11. Carver, C. S., and Scheier, M. F. (1982). Control theory: A useful conceptual framework for personality–social, clinical, and health psychology. Psychological Bulletin. 引用于 §38.2
  12. Coburn, A., Vartanian, O., Kenett, Y. N., Nadal, M., Hartung, F., Hayn-Leichsenring, G., … Chatterjee, A. (2020). Psychological and neural responses to architectural interiors. Cortex. 引用于 §38.11
  13. de Wit, S., Kindt, M., Knot, S. L., Verhoeven, A. A. C., Robbins, T. W., Gasull-Camos, J., … Gillan, C. M. (2018). Shifting the balance between goals and habits: Five failures in experimental habit induction. Journal of Experimental Psychology: General. 引用于 §38.2
  14. Deci, E. L., and Ryan, R. M. (2000). The "What" and "Why" of Goal Pursuits: Human Needs and the Self-Determination of Behavior. Psychological Inquiry. 引用于 §38.2
  15. Dosen, A. S., and Ostwald, M. J. (2016). Evidence for prospect-refuge theory: a meta-analysis of the findings of environmental preference research. City, Territory and Architecture. 引用于 §38.11
  16. Dubey, A., Naik, N., Parikh, D., Raskar, R., and Hidalgo, C. A. (2016). Deep Learning the City: Quantifying Urban Perception at a Global Scale. Computer Vision – ECCV 2016. 引用于 §38.11
  17. Eastwick, P. W., Sparks, J., Finkel, E. J., Meza, E. M., Adamkovič, M., Adu, P., … Coles, N. A. (2025). A worldwide test of the predictive validity of ideal partner preference matching. Journal of Personality and Social Psychology. 引用于 §38.7
  18. Falk, A., Becker, A., Dohmen, T., Enke, B., Huffman, D., and Sunde, U. (2018). Global Evidence on Economic Preferences*. The Quarterly Journal of Economics. 引用于 §38.6
  19. Forscher, P. S., Lai, C. K., Axt, J. R., Ebersole, C. R., Herman, M., Devine, P. G., and Nosek, B. A. (2019). A meta-analysis of procedures to change implicit measures. Journal of Personality and Social Psychology. 引用于 §38.1
  20. Franzen, A., and Mader, S. (2023). The power of social influence: A replication and extension of the Asch experiment. PLOS ONE. 引用于 §38.1 §38.13
  21. Fraser, I., Balcombe, K., Williams, L., and McSorley, E. (2021). Preference stability in discrete choice experiments. Some evidence using eye-tracking. Journal of Behavioral and Experimental Economics. 引用于 §38.3
  22. Frey, R., Pedroni, A., Mata, R., Rieskamp, J., and Hertwig, R. (2017). Risk preference shares the psychometric structure of major psychological traits. Science Advances. 引用于 §38.6
  23. Frömer, R., Dean Wolf, C. K., and Shenhav, A. (2019). Goal congruency dominates reward value in accounting for behavioral and neural correlates of value-based decision-making. Nature Communications. 引用于 §38.2
  24. Galak, J., and Redden, J. P. (2018). The Properties and Antecedents of Hedonic Decline. Annual Review of Psychology. 引用于 §38.4
  25. Gawronski, B., Morrison, M., Phills, C. E., and Galdi, S. (2017). Temporal Stability of Implicit and Explicit Measures: A Longitudinal Analysis. Personality and Social Psychology Bulletin. 引用于 §38.1
  26. Germine, L., Russell, R., Bronstad, P. M., Blokland, G. A., Smoller, J. W., Kwok, H., … Wilmer, J. B. (2015). Individual Aesthetic Preferences for Faces Are Shaped Mostly by Environments, Not Genes. Current Biology. 引用于 §38.8
  27. Glickman, M., and Sharot, T. (2025). How human–AI feedback loops alter human perceptual, emotional and social judgements. Nature Human Behaviour. doi:10.1038/s41562-024-02077-2. 引用于 §38.1
  28. Helzer, E. G., Fleeson, W., Furr, R. M., Meindl, P., and Barranti, M. (2017). Once a Utilitarian, Consistently a Utilitarian? Examining Principledness in Moral Judgment via the Robustness of Individual Differences. Journal of Personality. 引用于 §38.5
  29. Hoeffler, S., and Ariely, D. (1999). Constructing Stable Preferences: A Look Into Dimensions of Experience and Their Impact on Preference Stability. Journal of Consumer Psychology. 引用于 §38.3
  30. Hoeffler, S., Ariely, D., and West, P. (2006). Path dependent preferences: The role of early experience and biased search in preference development. Organizational Behavior and Human Decision Processes. 引用于 §38.3
  31. Iyengar, S. S., and Lepper, M. R. (2000). When choice is demotivating: Can one desire too much of a good thing? Journal of Personality and Social Psychology. 引用于 §38.3
  32. Journal of Marketing Research (2024). Expression of Concern:“The Dishonesty of Honest People: A Theory of Self-Concept Maintenance”. Journal of Marketing Research. 引用于 §38.5
  33. Keswani, V., Conitzer, V., Heidari, H., Borg, J. S., and Sinnott-Armstrong, W. (2024). On the Pros and Cons of Active Learning for Moral Preference Elicitation. AIES. 引用于 §38.5
  34. Keswani, V., Cousins, C., Nguyen, B., Conitzer, V., Heidari, H., Borg, J. S., and Sinnott-Armstrong, W. (2026). Moral Change or Noise? On Problems of Aligning AI With Temporally Unstable Human Feedback. AAAI. 引用于 §38.5 §38.13
  35. Kettlewell, N., Morris, R. W., Ho, N., Cobb-Clark, D. A., Cripps, S., and Glozier, N. (2020). The differential impact of major life events on cognitive and affective wellbeing. SSM - Population Health. 引用于 §38.4
  36. Kouchaki, M., and Smith, I. H. (2014). The Morning Morality Effect. Psychological Science. 引用于 §38.9
  37. Kurdi, B., Seitchik, A. E., Axt, J. R., Carroll, T. J., Karapetyan, A., Kaushik, N., … Banaji, M. R. (2019). Relationship between the Implicit Association Test and intergroup behavior: A meta-analysis. American Psychologist. 引用于 §38.1
  38. Lindqvist, E., Östling, R., and Cesarini, D. (2020). Long-Run Effects of Lottery Wealth on Psychological Well-Being. The Review of Economic Studies. 引用于 §38.4
  39. Mazar, N., Amir, O., and Ariely, D. (2008). The Dishonesty of Honest People: A Theory of Self-Concept Maintenance. Journal of Marketing Research. 引用于 §38.5
  40. McShane, B. B., and Böckenholt, U. (2018). Multilevel Multivariate Meta-analysis with Application to Choice Overload. Psychometrika. 引用于 §38.3
  41. Monin, B., and Miller, D. T. (2001). Moral credentials and the expression of prejudice. Journal of Personality and Social Psychology. 引用于 §38.5
  42. Rehren, P., and Sinnott-Armstrong, W. (2022). How Stable are Moral Judgments? Review of Philosophy and Psychology. doi:10.1007/s13164-022-00649-7. 引用于 §38.5
  43. Reiter, A. M. F., Moutoussis, M., Vanes, L., Kievit, R., Bullmore, E. T., Goodyer, I. M., … Dolan, R. J. (2021). Preference uncertainty accounts for developmental effects on susceptibility to peer influence in adolescence. Nature Communications. 引用于 §38.1
  44. Reutskaja, E., Lindner, A., Nagel, R., Andersen, R. A., and Camerer, C. F. (2018). Choice overload reduces neural signatures of choice set value in dorsal striatum and anterior cingulate cortex. Nature Human Behaviour. 引用于 §38.3
  45. Roberts, B. W., and DelVecchio, W. F. (2000). The rank-order consistency of personality traits from childhood to old age: A quantitative review of longitudinal studies. Psychological Bulletin. 引用于 §38.6
  46. Rotella, A., Jung, J., Chinn, C., and Barclay, P. (2026). Observation Moderates the Moral Licensing Effect: A Meta-Analytic Test of Interpersonal and Intrapsychic Mechanisms. Personality and Social Psychology Bulletin. 引用于 §38.5 §38.13
  47. Ruttan, R. L., and Nordgren, L. F. (2021). Instrumental use erodes sacred values. Journal of Personality and Social Psychology. 引用于 §38.5
  48. Scheibehenne, B., Greifeneder, R., and Todd, P. M. (2010). Can There Ever Be Too Many Options? A Meta-Analytic Review of Choice Overload. Journal of Consumer Research. 引用于 §38.3
  49. Schwartz, S. H. (1992). Universals in the Content and Structure of Values: Theoretical Advances and Empirical Tests in 20 Countries. Advances in Experimental Social Psychology. 引用于 §38.2
  50. Schwarz, N., and Clore, G. L. (1983). Mood, misattribution, and judgments of well-being: Informative and directive functions of affective states. Journal of Personality and Social Psychology. 引用于 §38.4
  51. Shu, L. L., Mazar, N., Gino, F., Ariely, D., and Bazerman, M. H. (2012). RETRACTED: Signing at the beginning makes ethics salient and decreases dishonest self-reports in comparison to signing at the end. Proceedings of the National Academy of Sciences. 引用于 §38.5
  52. Silver, A. M., Stahl, A. E., Loiotile, R., Smith-Flores, A. S., and Feigenson, L. (2020). When Not Choosing Leads to Not Liking: Choice-Induced Preference in Infancy. Psychological Science. 引用于 §38.1 §38.13
  53. Smeele, N. V., van Cranenburgh, S., Donkers, B., Schermer, M. H., and de Bekker-Grob, E. W. (2025). Taboo trade-off aversion in choice behaviors: A discrete choice model and application to health-related decisions. Social Science & Medicine. 引用于 §38.5
  54. Tetlock, P. E., Kristel, O. V., Elson, S. B., Green, M. C., and Lerner, J. S. (2000). The psychology of the unthinkable: Taboo trade-offs, forbidden base rates, and heretical counterfactuals. Journal of Personality and Social Psychology. 引用于 §38.5
  55. Vaidis, D. C., Sleegers, W. W. A., van Leeuwen, F., DeMarree, K. G., Sætrevik, B., Ross, R. M., … Priolo, D. (2024). A Multilab Replication of the Induced-Compliance Paradigm of Cognitive Dissonance. Advances in Methods and Practices in Psychological Science. 引用于 §38.1 §38.13
  56. Vecchione, M., Schwartz, S. H., Davidov, E., Cieciuch, J., Alessandri, G., and Marsicano, G. (2020). Stability and change of basic personal values in early adolescence: A 2‐year longitudinal study. Journal of Personality. 引用于 §38.2
  57. Verschuere, B., Meijer, E. H., Jim, A., Hoogesteyn, K., Orthey, R., McCarthy, R. J., … Yıldız, E. (2018). Registered Replication Report on Mazar, Amir, and Ariely (2008). Advances in Methods and Practices in Psychological Science. 引用于 §38.5 §38.13
  58. Walter, K. V., Conroy-Beam, D., Buss, D. M., Asao, K., Sorokowska, A., Sorokowski, P., … Zupančič, M. (2020). Sex Differences in Mate Preferences Across 45 Countries: A Large-Scale Replication. Psychological Science. 引用于 §38.7
  59. Wilson, T. D., Lindsey, S., and Schooler, T. Y. (2000). A model of dual attitudes. Psychological Review. 引用于 §38.1
  60. Wood, W., Mazar, A., and Neal, D. T. (2022). Habits and Goals in Human Behavior: Separate but Interacting Systems. Perspectives on Psychological Science. 引用于 §38.2
  61. Xiao, Q., Li, L. C., Au, Y. L., Tan, S. N., Chung, W. T., and Feldman, G. (2024). Licensing via Credentials: Replication Registered Report of Monin and Miller (2001) with Extensions Investigating the Domain-Specificity of Moral Credentials and the Association Between the Credential Effect and Trait Reputational Concern. International Review of Social Psychology. 引用于 §38.5
  62. Yap, S. C. Y., Wortman, J., Anusic, I., Baker, S. G., Scherer, L. D., Donnellan, M. B., and Lucas, R. E. (2017). The effect of mood on judgments of subjective well-being: Nine tests of the judgment model. Journal of Personality and Social Psychology. 引用于 §38.4 §38.13
  63. Zickfeld, J., Gonzalez, A. S. R., and Mitkidis, P. (2024). Investigating the Morning Morality Effect and its Mediating and Moderating Factors. PsyArXiv. 预印本引用于 §38.9