三思而行——第一章(外部视角:为什么大布朗是个糟糕的赌注)
经哈佛商学院出版社许可重印。内容摘自迈克尔·莫布森的《三思而后行:善用反直觉的力量》一书。© 2009 迈克尔·J·莫布森 版权所有。
Reprinted by permission of Harvard Business Press. Excerpted from Think Twice: Harnessing the Power of Counterintuition by Michael Mauboussin. Copyright © 2009 by Michael J. Mauboussin; all rights reserved.
CHAPTER ON E
CHAPTER ON E
外部视角:为什么大哥大·布朗是个糟糕的赌注
The Outside View Why Big Brown Was a Bad Bet
“这事十拿九稳了。”里克·达特罗如此断言他的赛马“大布朗”在 2008 年有望夺得众人垂涎的三冠王。赢取三冠王是一项了不起的壮举。一匹马必须在短短五周内,在三个不同长度的赛道上,接连赢下肯塔基德比、普瑞克内斯锦标赛和贝尔蒙特锦标赛。在“大布朗”挑战之前,过去一个世纪里只有十一匹马成功过,而此前的三十年里一匹都没有。当时“大布朗”距离赛马界的“不朽传奇”只差一场比赛了。¹ 作为驯马师,达特罗有理由乐观。这匹三岁小公马不仅前五战全胜,而且每一场都赢得无可争议。尽管赌马赔率只给了它 25% 的概率赢下肯塔基德比,“大布朗”却以领先四个又四分之三马身的优势获胜。它在普瑞克内斯锦标赛上表现更加强势,以领先五个又四分之一马身的长度冲过终点线——即便它的骑师在最后直道已经放松了缰绳。在最后一场贝尔蒙特锦标赛中,“大布朗”面对的是水平平平的对手,它最大的挑战者“赌场之旅”还在最后一刻退出了比赛。
“I t’S A FORE GONE CONCL USION . ” So proclaimed Rick Dutrow on the likelihood that his race horse, Big Brown, would capture the coveted Triple Crown in 2008. Winning the Triple Crown is a tremendous feat. A horse must win the Kentucky Derby, the Preakness Stakes, and the Belmont Stakes on three tracks of different lengths over just five weeks. Before Big Brown’s attempt, only eleven horses had succeeded in the preceding century, and none had done so in the previous thirty years. Here was Big Brown, only one race away from horse-racing “immortality.”1 Dutrow, the horse’s trainer, had reason to be optimistic. Not only was his three-year-old colt undefeated in his first five starts, he was dominant. Although the odds makers placed only a 25 percent probability on his winning the Kentucky Derby, Big Brown won by four-and-three-quarters lengths. He was stronger still in the Preakness, crossing the finish line five-and-one-quarter lengths ahead of the field, even though his jockey eased him coming down the home stretch. In his last race, the Belmont, Big Brown faced mediocre competition, and his biggest challenger, Casino Drive, dropped out of the race at the last minute.
不出所料,市场对“大棕”的热情不断升温。嗅到机会的 UPS——正是“大棕”这个名字所效仿的公司——签约了
Not surprisingly, enthusiasm for Big Brown built. Sensing opportunity, UPS, the company after which Big Brown was named, signed
一笔营销合约,内容是把一条公司标志印在“大布朗”领骑员的外套上。大多数赛马专业人士都押它能赢。再看“大布朗”本身——它被描绘得强壮、自信、准备就绪。驯马师达彻滔滔不绝地说:“它看起来简直不能再好了。我完全找不出“大布朗”身上有任何瑕疵。我看到的是最完美的画面。我信心十足,简直难以置信。”2 马迷们也深有同感:尽管酷暑难耐,但这场关键比赛的观赛人数竟比前一年翻了一番——人们渴望亲眼见证历史。
a marketing deal that included a corporate logo on the jacket of Big Brown’s outrider. A majority of racetrack pros picked him to win the race. And then there was Big Brown himself. He was por-trayed as strong, confident, and ready. Dutrow gushed, “He looks as good as he can possibly look. I can’t find any flaws whatsoever in Big Brown. I see the prettiest picture. I’m so confident, it’s unbelievable.”2 The fans agreed: attendance for the pivotal race was double what it had been in the previous year despite the sweltering heat, as the crowd yearned to see history made.
大布朗确实创造了历史,没错。只是这历史与所有人预想的都不一样。他跑了个倒数第一——这在所有三冠王竞争者中是绝无仅有的。赛后兽医对大布朗进行了全面体检,结果显示它一切正常。它那反复无常的表现,让人想起实验室研究者所说的“哈佛法则”:“在最严格控制的压强、温度、容积、湿度及其他变量的条件下,生物体还是会随心所欲地行事。”然而,还有一种看待大布朗赢得三冠王几率的视角,这个视角对其跻身赛马神殿的前景要悲观得多。
Big Brown made history, all right. It just wasn’t the kind of history everyone expected. He finished dead last, which no Triple Crown contender had ever done.3 Veterinarians gave Big Brown a full physical exam following the race and he appeared to be fine. His capricious performance evoked what lab researchers call Harvard’s Law, “Under the most rigorously controlled conditions of pressure, temperature, volume, humidity, and other variables, the organism will do as it damn well pleases.”4 However, there was another way of looking at Big Brown’s chances of winning the Triple Crown, one that was far less optimistic about his prospects of joining the pantheon of horse racing.
这一观点提出了一个简单的问题:其他马匹在处于“大布朗”的位置时,表现有多成功?
This point of view asked a simple question: how successful were other horses when they were in Big Brown’s position?
史蒂文·克里斯特是一位才华横溢的作家和著名赛马评注人,他提供了一些令人警醒的数据。在赢得肯塔基德比和普瑞克内斯锦标赛后有机会冲击三冠王的 29 匹马中,只有 11 匹成功,成功率不到 40%。但仔细审视这些数据后,1950 年前后呈现出极其鲜明的反差。1950 年之前,尝试赢得三冠王的 9 匹马中有 8 匹成功。1950 年之后,20 匹马中只有 3 匹获胜。很难解释为何成功率从将近 90% 骤降至仅 15%,但合乎逻辑的因素包括更优良的育种(培育出更多优质马驹)和更大的出赛马群规模。
Steven Crist, a talented writer and renowned handicapper, pro-vided some sobering statistics.5 Of the twenty-nine horses with a chance to capture the Triple Crown after winning the Kentucky Derby and the Preakness Stakes, only eleven triumphed, a success rate less than 40 percent. But a closer examination of those statistics yielded a stark difference before and after 1950. Before 1950, eight of the nine horses attempting to win the Triple Crown succeeded. After 1950, only three of twenty horses won. It’s hard to know why the achievement rate dropped from nearly 90 percent to just 15 percent, but logical factors include better breeding (leading to more quality foals) and bigger starting fields.
虽然 15% 的成功率可能会引发一些担忧,但这并未考虑“大棕”的天生能力。
While a 15 percent rate of success may raise some concern, it doesn’t take into consideration Big Brown’s innate ability and
令人瞩目的战绩。毕竟,并非所有有望赢得三冠王的赛马都具备同样的天赋。比较赛马的一种方法是贝耶速度指数,该指数根据比赛时间、赛道速度以及天气条件,为马匹的表现赋予一个数字。速度指数越高越好。
impressive track record. After all, not all of the horses in a position to win the Triple Crown had similar talent. One way to compare horses is the Beyer Speed Figure, which assigns a number to a horse’s performance based on the time of the race and the speed of the track, given the weather conditions. Higher speed figures are better.
表 1-1 显示了过去七匹三冠王挑战马(包括大布朗)在头两场三冠赛事中的速度数据。样本较小,因为速度数据从 1991 年起才广泛可得。虽然骑师的动作可能让它的普瑞克尼斯赛数据降低了几个点,但与其他马相比,大布朗看起来简直就是铁蹄慢吞吞。即使考虑到贝蒙锦标赛的阵容实力一般,大布朗也显然并非稳赢。然而,投注者将大布朗的赔率推到了令人狂热的 3 赔 10,意味着他们认为它赢得最后一关的概率超过 75%。克里斯和其他精明的赌马人则有马感,能看出投注板大幅高估了大布朗的获胜概率。
Table 1-1 shows the speed figures in the first two Triple Crown races for the last seven aspirants, including Big Brown. The sample is small because speed figures have been widely available only since 1991. While his jockey’s actions likely pared a few points from his Preakness figure, Big Brown looked downright lead-hoofed when compared to the other horses. Even considering the so-so Belmont field, it was obvious that Big Brown was not certain to win. Yet the bettors had Big Brown’s odds at a euphoric three-to-ten, implying he had more than a 75 percent probability of winning the final leg. Crist and other sharp handicappers had the horse sense to recognize the tote board substantially overstated Big Brown’s chance of winning.
这两种对立观点的对比,暴露了我们的第一个错误——倾向于偏爱内部视角而非外部视角。⁶ 内部视角在考虑问题时,会聚焦于具体任务,利用手头直接可用的信息,并基于这些信息做出预测。
These contrasting points of view reveal our first mistake, a ten-dency to favor the inside view over the outside view.6 An inside view considers a problem by focusing on the specific task and by using information that is close at hand, and makes predictions based on
TABLE 1-1
TABLE 1-1
三冠王候选赛驹的贝耶速度指数
Beyer Speed Figures for Triple Crown contenders
| 马匹 | 肯塔基德比 | 普利克内斯锦标 | 总分 |
|---|---|---|---|
| 银魅 | 115 | 118 | 233 |
| 机灵鬼 | 107 | 118 | 225 |
| 滑稽戏 | 109 | 114 | 223 |
| 战争标志 | 114 | 109 | 223 |
| 真静 | 107 | 111 | 218 |
| 魅力型 | 108 | 107 | 215 |
| 大棕 | 109 | 100 | 209 |
Horse Kentucky Derby Preakness Total Silver Charm 115 118 233 Smarty Jones 107 118 225 Funny Cide 109 114 223 War Emblem 114 109 223 Real Quiet 107 111 218 Charismatic 108 107 215 Big Brown 109 100 209
来源:史蒂文·克里斯(Steven Crist)
Source: Steven Crist.
那套狭窄且独特的输入数据,这些输入可能包含轶事证据和谬误感知。大多数人构建未来模型时都会采用这种方法,事实上所有形式的规划都普遍如此。里克·达特罗(Rick Dutrow)和“大布朗”的其他粉丝们基本只看内部视角,包括这匹马赢得比赛的战绩和威风凛凛的外表。这很自然,但几乎总是描绘出一幅过于乐观的画面。
that narrow and unique set of inputs. These inputs may include anecdotal evidence and fallacious perceptions. This is the approach that most people use in building models of the future and is indeed common for all forms of planning. Rick Dutrow and the other fans of Big Brown dwelled largely on the inside view, including the horse’s wins and imposing physical appearance. This comes natu-rally but almost always paints too optimistic a picture.
外部视角要求我们寻找是否存在类似情境,从而为决策提供统计依据。它不把问题看作独一无二的个案,而是想知道别人是否面临过类似问题,如果答案是肯定的,结果如何。外部视角是一种违背直觉的思维方式。因为它迫使人们抛开自己收集的所有宝贵信息。使用外部视角的马彩预测者认为,“大布朗”的投注价值极低——根据其他处于同样处境马匹的经验,它获胜的概率远低于博彩赔率牌上的显示。外部视角常常能为决策者提供非常宝贵的现实检验。
The outside view asks if there are similar situations that can provide a statistical basis for making a decision. Rather than seeing a problem as unique, the outside view wants to know if others have faced comparable problems and, if so, what happened. The outside view is an unnatural way to think, precisely because it forces people to set aside all the cherished information they have gathered. Handicappers using the outside view judged Big Brown to be a very poor bet, as the experience of other horses in the same spot suggested a probability of winning that was much lower than what was on the tote board. The outside view can often create a very valuable reality check for decision makers.
为什么人们倾向于采用内部视角?我们大多数人在很多时候都过于乐观。社会心理学家区分了三种导致人们采用内部视角的幻觉⁷。要介绍第一种幻觉,请花点时间(诚实地!)回答以下问题,用“是”或“否”:
Why do people tend to embrace the inside view? Most of us are unduly optimistic a good deal of the time. Social psychologists dis-tinguish three illusions that lead people to the inside view.7 To introduce the first illusion, take a moment to answer (hon-estly!) the following questions either yes or no:
• 我是一个高于平均水平的司机。
• I am an above-average driver.
• 我有高于平均水平的幽默鉴赏力。
• I have an above-average ability to judge humor.
• 我的职业表现让我排在公司前一半的水平。
• My professional performance places me in the top half of my organization.
如果你和大多数人一样,那么你对这三个问题的回答都是“是”。
If you are like most people, you said yes to all three questions.
这项调查展现了优越感错觉,即人们对自己抱有脱离现实的正面评价。当然,不可能人人都高于平均水平。在 1976 年的一项经典调查中,美国大学理事会(College Board)请高中考生在一系列指标上给自己打分——
This shows the illusion of superiority, which suggests people have an unrealistically positive view of themselves. Of course, not everyone can be above average. In a classic 1976 survey, the College Board asked high school test takers to rate themselves on a host of crite-
85% 的人认为自己高于中位数水平。
ria. Eighty-five percent considered themselves above the median in
与他人相处的能力方面,比中位数高 70%;领导他人的能力方面,比中位数高 60%;体育方面也高出 60%。一项调查显示,超过 80% 的人认为自己比半数司机驾驶技术更娴熟。⁸ 值得注意的是,能力最差的人,往往认为自己能达成的与实际情况之间的差距最大。⁹ 在一项研究中,研究人员要求受试者对自己在语法测试中的感知能力和可能的成功程度进行评分。图 1-1 显示,表现最差的受试者严重高估了自己的能力,他们以为自己能排在倒数第二高的四分位,实际结果却落在了最末的四分位。此外,即使人们承认自己低于平均水平,他们也倾向于将自己的缺点视为无关紧要而敷衍过去。
getting along with others, 70 percent above the median in ability to lead others, and 60 percent above the median in sports. One survey showed that more than 80 percent of people believed that they were more skillful than half of all drivers.8 Remarkably, the least capable people often have the largest gaps between what they think they can do and what they actually achieve.9 In one study, researchers asked subjects to rate their perceived ability and likely success on a grammar test. Figure 1-1 shows that the poorest performers dramatically overstated their ability, thinking that they would be in the next-to-highest quartile. They turned in results in the bottom quartile. Furthermore, even when individuals do acknowledge that they are below average, they tend to dismiss their shortcomings as inconsequential.
FIGURE 1-1
FIGURE 1-1
能力最差的人往往最自负。
The least competent are often the most confident
100 Perceived ability
100 Perceived ability
90 感知测试分数
90 Perceived test score
| 实际考试成绩 | |||
| 80 | |||
| 70 | |||
| 60 | |||
| 百分位 | 50 | ||
| 40 | |||
| 30 | |||
| 20 | |||
| 10 | |||
| 0 | |||
| 底层 | 第二四分位 | 第三四分位 | 顶层 |
| 四分位 | 四分位 | 四分位 | 四分位 |
Actual test score 80 70 60 Percentile 50 40 30 20 10 0 Bottom 2nd 3rd Top quartile quartile quartile quartile
来源:贾斯汀·克鲁格与戴维·邓宁,《无能且不自知:为何无法识别自身无能会导致自我评价膨胀》,《人格与社会心理学杂志》第 77 期
Source: Justin Kruger and David Dunning, “Unskilled and Unaware of It: How Difficulties in Recognizing One’s Own Incompetence Lead to Inflated Self-Assessments.” Journal of Personality and Social Psychology 77,
no. 6 (1999): 1121–1134.
no. 6 (1999): 1121–1134.
第二个是乐观幻觉。大多数人对自己的未来比他人都要乐观。举例来说,研究人员让大学生估测自己在生活中遇到各种好与坏经历的概率。这些学生判断自己遇到好经历的可能性远高于同龄人,遇到坏经历的可能性则远低于同龄人。¹⁰ 最后是控制幻觉。人们表现得仿佛随机事件能被自己掌控。比如,掷骰子的人在想要掷出小点数时会轻轻掷,想掷出大点数时则用力掷。在一项研究中,研究人员让两组办公室职员参与一个彩票活动,费用 1 美元,奖金 50 美元。其中一组可以自己选择彩票卡,另一组则没有选择权。当然,中奖的概率由运气决定,但职员们的表现并非如此。
The second is the illusion of optimism. Most people see their future as brighter than that of others. For example, researchers asked college students to estimate their chances of having various good and bad experiences during their lives. The students judged themselves far more likely to have good experiences than their peers, and far less likely to have bad experiences.10 Finally, there is the illusion of control. People behave as if chance events are subject to their control. For instance, people roll-ing dice throw softly when they want to roll low numbers and hard for high numbers. In one study, researchers asked two groups of office workers to participate in a lottery, with a $1 cost and a $50 prize. One group was allowed to choose their lottery cards, while the other group had no choice. Luck determined the probability of winning, of course, but that’s not how the workers behaved.
抽奖开始前,一位研究人员询问参与者愿意以什么价格出售自己的卡片。获准自行选卡的那组人,平均报价接近 9 美元;而未被允许选卡的那组人,报价则低于 2 美元。那些相信自己拥有一定控制权的人,会觉得自己成功的概率比实际更高。
Before the drawing, one of the researchers asked the partici-pants at what price they would be willing to sell their cards. The mean offer for the group that was allowed to choose cards was close to $9, while the offer from the group that had not chosen was less than $2. People who believe that they have some control have the perception that their odds of success are better than they actually
。没有掌控感的人不会经历同样的偏差。¹¹ 我必须承认,我的职业——主动资金管理——可能是专业领域中体现“掌控幻觉”的最佳例子之一。研究人员已经表明,从整体来看,主动构建投资组合的资金经理长期来看提供的回报低于市场指数,这一发现每家投资公司都承认。¹² 原因相当直接:市场竞争高度激烈,而资金经理收取的费用会侵蚀回报。市场还有相当程度的随机性,确保所有投资者都会时而看到好业绩、时而看到差业绩。尽管有这些证据,主动资金经理的行动却仿佛他们能够战胜概率、提供超越市场的回报。这些投资公司依赖内部视角来证明自己的策略和收费合理性。
are. People who don’t have a sense of control don’t experience the same bias.11 I must concede that my occupation, active money management, may be one of the best examples of the illusion of control in the professional world. Researchers have shown that, in aggregate, money managers who actively build portfolios deliver returns lower than the market indexes over time, a finding that every investment firm acknowledges.12 The reason is pretty straightforward: markets are highly competitive, and money managers charge fees that diminish returns. Markets also have a good dose of randomness, assuring that all investors see good and poor results from time to time. Despite the evidence, active money managers behave as if they can defy the odds and deliver market-beating returns. These investment firms rely on the inside view to justify their strategies and fees.
成功的几率微乎其微……但我除外
The Odds of Success Are Poor . . . But Not for Me
大量专业人士普遍依赖内部视角来做重要决策,结果可想而知地糟糕。这不是说这些决策者玩忽职守、天真幼稚或心怀恶意。
A vast range of professionals commonly lean on the inside view to make important decisions with predictably poor results. This is not to say these decision makers are negligent, naïve, or malicious.
受到这三种错觉的鼓舞,大多数人相信自己做出了正确的决定,并且坚信结果会令人满意。既然你已经意识到内部视角与外部视角的区别,你就可以更仔细地衡量自己和他人的决策。让我们来看几个例子。
Encouraged by the three illusions, most believe they are making the right decision and have faith that the outcomes will be satisfac-tory. Now that you are aware of the distinction between the inside and outside view, you can measure your decisions and the decisions of others more carefully. Let’s look at some examples.
企业并购(M&A)是一项全球每年高达数万亿美元的大生意。企业花费巨额资金识别、收购和整合公司,以期获得战略优势。毫无疑问,公司做交易时都怀着最良好的初衷。
Corporate mergers and acquisitions (M&A) are a multitrillion dollar global business year in and year out. Corporations spend vast sums identifying, acquiring, and integrating companies in order gain a strategic edge. There is little doubt that companies make deals with the best of intentions.
问题在于,多数交易并没有为收购公司的股东创造价值(被收购公司的股东通常表现不错)。事实上,研究者估算,当一家公司收购另一家公司时,收购方的股票约有三分之二的时间是下跌的¹³。考虑到多数经理人的明确目标是提升价值——而且他们的薪酬通常与股价挂钩——并购市场的活跃程度显得有些令人意外。
The problem is that most deals don’t create value for the shareholders of the acquiring company (shareholders of the companies that are bought do fine, on average). In fact, researchers estimate that when one company buys another, the acquiring company’s stock goes down roughly two-thirds of the time.13 Given that most managers have an explicit objective of increas-ing value—and that their compensation is often tied to the stock price—the vigor of the M&A market appears moderately surpris-
这里的解释是,尽管多数高管都清楚并购的整体记录并不理想,但他们相信自己能战胜概率。
ing. The explanation is that while most executives recognize that the overall M&A record is not good, they believe that they can beat the odds.
“顶级海滨地产”——这是 2008 年 7 月陶氏化学(Dow Chemical)同意收购罗门哈斯(Rohm and Haas)后,其首席执行官对该公司的形容。陶氏对竞价战毫无惧色,尽管这场争夺战已将它必须支付的溢价推至高达 74%。相反,这位首席执行官宣称,这笔交易是“陶氏朝着成为盈利增长公司迈出的决定性一步”。陶氏管理层的热情,处处透着内部视角的典型特征。交易宣布时,陶氏化学的股价大幅下跌。
“A high-quality beachfront property” is how the chief execu-tive officer of Dow Chemical described Rohm and Haas after Dow agreed to acquire the company in July 2008. Dow was undaunted by the bidding war, which had driven the price premium it had to pay to a steep 74 percent. Instead, the CEO declared the deal “a decisive step towards establishing Dow as an earnings-growth company.”14 The enthusiasm of Dow’s management had all the hallmarks of the inside view. When the deal was announced, the stock price of Dow Chemical slumped
4%,使得这笔交易登上了不断增长的收购亏损榜榜首。
4 percent, putting the deal on top of a growing pile of losses suf-fered through acquisitions.
基础数学解释了为什么大多数公司在收购另一家企业时不会创造价值。对收购方而言,价值的变化等于两家公司合并后现金流增加额(协同效应)与收购方支付的超出市场价值的溢价两者之间的差额。公司总想得到的比付出的多。因此,如果协同效应超过溢价,收购方的股价就会上涨;反之则下跌。在此案例中,根据陶氏自身的数据,协同效应的价值低于其支付的溢价,这恰恰证明了股价下跌的合理性。抛开华丽的辞藻不谈,这些数字对陶氏化学的股东而言并不有利。
Basic math explains why most companies don’t add value when they acquire another firm. The change in value for the buyer equals the difference between the increase in cash flow from combining the two companies (synergies) and the amount over the market value that the acquirer pays (premium). Companies want to get more than they pay for. So if synergies exceed the premium, the price of the buyer’s stock will rise. If not, it will fall. In this case, the value of the synergy—based on Dow’s own figures—was less than the premium it paid, justifying a drop in price. Glowing rhetoric aside, the numbers were not good for the shareholders of Dow Chemical.15
“多个轶事加在一起,也并不等于证据。”
The Plural of Anecdote Is Not Evidence
几年前,我父亲被诊断为晚期癌症。化疗失败后,他基本没有了选择。有一天,他打电话来征求我的意见。他在一本杂志上看到一则关于某种替代癌症疗法的广告,宣称效果近乎奇迹,并指向一个满是好评推荐的网站。如果他把资料发给我,我能不能告诉他我的看法?
A few years ago, my father was diagnosed with late-stage cancer. After the chemotherapy failed, he was basically out of options. One day, he called seeking my advice. He had read a magazine advertisement about an alternative cancer treatment that claimed near-miraculous results and pointed to a Web site full of glowing testimonials. If he sent me the information, would I tell him what I thought?
研究没花多长时间。没有任何设计严谨的研究证明过该疗法的疗效,支持这种方法的证据不过是一堆个人轶事。父亲回电话时,我从他的语气里听出,他已经拿定了主意。
It didn’t take long to do the research. No well-constructed studies had shown the treatment’s efficacy, and the evidence in favor of the approach amounted to a collection of anecdotes. When my father called back, I could hear in his voice that his mind was made
挂了电话,我心里很矛盾。我希望相信那个故事,跟随内心的视角。我希望父亲能好起来。但我内心那个科学家告诫我,要坚持外部视角。即便考虑到安慰剂效应的强大力量,希望也不是一种策略。
up. Despite the substantial cost and taxing travel, he wanted to pursue this long-shot alternative. When he asked me what I thought, I told him, “I try to think like a scientist. And based on everything I can see, this won’t work.” Hanging up the phone, I felt torn. I wanted to believe the story and go with the inside view. I wanted my father to be well again. But the scientist in me admonished me to stick with the outside view. Even considering the power of the placebo effect, hope is not a strategy.
那次事件之后不久,我父亲就去世了,但这段经历促使我思考,我们究竟是如何决定自己的医疗方案的。
My father died shortly after that episode, but the experience com-pelled me to think about how we decide about our medical treatments.
长期以来,家长式模式主导着医生与患者之间的关系。医生诊断病情后,选择他们认为对患者最合适的治疗方案。如今的患者了解信息更多,通常希望参与决策过程。医生和患者经常讨论各种治疗方案的利弊,共同选择最佳行动方案。确实,研究表明,参与决策的患者对医疗治疗的满意度更高。
For a long time, the paternalistic model reigned in relationships between physicians and patients. Physicians would diagnose a condi-tion and select the treatment that seemed best for the patient. Patients nowadays are more informed and generally want to take part in making decisions. Physicians and patients frequently discuss the pros and cons of various treatments and together select the best course of action. Indeed, studies show that patients involved in making those decisions are more satisfied with their medical treatment.
但研究也表明,患者经常做出不符合自身最佳利益的选择,原因往往在于未能考虑外部视角。16 在一项研究中,研究人员向受试者提供了一种虚构疾病以及多种治疗方案。每位受试者要从两种治疗方案中做出选择。第一种是对照治疗,有效性为 50%。第二种则包含一个虚构患者的正面、中性或负面案例,以及四种可能的有效概率(从 30% 到 90%),总共构成十二种选项。
But research also suggests that patients regularly make choices that are not in their best interests, often due to a failure to consider the outside view.16 In one study, researchers presented subjects with a fictitious disease and various treatments. Each subject had a choice between two treatments. The first, the control treatment, had 50 percent effectiveness. The second was one of twelve options that combined a positive, neutral, or negative anecdote about a fic-tional patient with four possible levels of effectiveness, ranging from 30 percent to 90 percent.
这些故事产生了巨大影响,在决策过程中完全压倒了基础概率数据。表 1-2 说明了这一点:当治疗方案搭配一个失败案例的故事时,患者选择该方案(该方案本身有 90% 的有效率)的比例不足 40%。
The stories made a huge difference and swamped the base-rate data in the decision-making process. Table 1-2 tells the tale. Patients selected a treatment with 90 percent effectiveness less than 40 percent of the time when it was paired with a story about a failed TABLE 1-2
轶事比解药更重要吗?
Are anecdotes more important than antidotes?
选择治疗方案的受试者百分比
Percent of subjects choosing the treatment
| 基础比率 | ||||
|---|---|---|---|---|
| 90% | 70% | 50% | 30% | |
| 正面轶事 | 88 | 92 | 93 | 78 |
| 中性轶事 | 81 | 81 | 69 | 29 |
| 负面轶事 | 39 | 43 | 15 | 7 |
BASE RATE 90% 70% 50% 30% Positive anecdote 88 92 93 78 Neutral anecdote 81 81 69 29 Negative anecdote 39 43 15 7
来源:Angela K. Freymuth 与 George F. Ronan,《模拟患者决策:基础率与轶事信息的作用》,《临床心理学在医疗环境中的应用》期刊,第 11 卷,第 3 期(2004 年):第 211–216 页。
Source: Angela K. Freymuth and George F. Ronan, “Modeling Patient Decision-Making: The Role of Base-Rate and Anecdotal Information,” Journal of Clinical Psychology in Medical Settings 11, no. 3 (2004): 211–216.
而与之相反的是,当近 80% 的患者看到一个成功案例后,他们选择了一种仅 30% 有效的治疗方案。这项研究的结果与我父亲的行为完全一致。
patient. Conversely, nearly 80 percent of the patients selected a treatment with 30 percent effectiveness when it was matched with a success story. The results of this study were fully consistent with my father’s behavior.
尽管患者了解情况并积极参与是好事,但他们面临的风险是,可能会受到主要依赖个人经历的来源的影响,比如朋友、家人、互联网和大众媒体。医生可能会发现,用个人经历向患者传达自己的观点是一种有效方式。但医生和患者都应注意,不要忽视科学证据。¹⁷
While it’s good for patients to be informed and engaged, they run the risk of being influenced by sources that rely predominantly on anecdotes, including friends, family, the Internet, and mass media. Doctors might find anecdotes to be an effective way of getting their points across to patients. But doctors and patients should be care-ful not to lose sight of the scientific evidence.17
按时且在预算内——也许下次吧
On Time and Within Budget—Maybe Next Time
如果你曾经参与过某个项目——不管是翻修房子、推出新产品还是赶工作截止日期——你应该对这个例子不陌生。人们很难准确预估一项工作需要多长时间、花多少钱。一旦判断出错,他们往往是低估了时间和成本。
You will be familiar with this example if you have ever been part of a project, whether it involved renovating a house, introducing a new product, or meeting a work deadline. People find it hard to estimate how long a job will take and how much it will cost. When they are wrong, they usually underestimate the time and expense.
心理学家称之为计划谬误。这里仍然是内部视角占主导,因为大多数人想象的是自己将如何完成这项任务。大约只有四分之一的人在制定计划时间表时,会参考来自自身经验或他人经验的基础比率数据。
Psychologists call this the planning fallacy. Here again, the inside view takes over, as the majority of people imagine how they will complete the task. Only about one-quarter of the population incor-porates the base-rate data either from their own experience or from that of others, while laying out planning timetables.
威尔弗里德·劳里埃大学心理学教授罗杰·比勒做过一个实验,说明了这一点。比勒和他的合作者让大学生估算完成一项学校作业所需的时间,给出了三个概率等级:50%、75% 和 99%。例如,一名受试者可能会说,他有 50% 的概率在下周一前完成作业,75% 的概率在周三前完成,99% 的概率在周五前完成。
Roger Buehler, a professor of psychology at Wilfrid Laurier University, did an experiment that illustrates the point. Buehler and his collaborators asked college students how long it would take to complete a school assignment with three levels of chance: 50, 75, and 99 percent. For example, a subject might say that there was a 50 percent chance that he would finish the project by next Monday, a 75 percent chance he’d be done by Wednesday, and a 99 percent chance by Friday.
图 1-2 展示了这些预估的准确度如何:当学生们自认为有 50% 概率完成的截止日期到来时,实际上只有 13% 的人交了作业。
Figure 1-2 shows how accurate the estimates were: when the deadline arrived for which the students had given themselves a 50 percent chance of finishing, only 13 percent actually turned in their
工作。当学生认为有 75% 的几率完成任务时,实际只有 19% 的人完成了项目。
work. At the point when the students thought there was 75 percent chance they’d be done, just 19 percent had completed the project.
所有学生都几乎确信他们能在截止日期前完成。但结果只有 45% 的人是对的。正如比勒及其合作研究者所言:“即便被要求做出一个高度保守的预测——一个他们感到几乎肯定会实现的预测——学生们在时间估算上的自信也远远超出了他们的实际成果。”¹⁸ 这项研究有一个有趣的转折:虽然人们在预测自己何时能完成任务方面糟糕透顶,但在预测别人时却相当准确。事实上,规划谬误体现了一个更广泛的原则。当人们被迫审视类似情况并看到成功的频率时,他们倾向于做出更准确的预测。如果你想知道某件事会如何……
All the students were virtually sure they’d be done by the final date. But only 45 percent turned out to be right. As Buehler and his fel-low researchers note, “Even when asked to make a highly conserva-tive forecast, a prediction that they felt virtually certain that they would fulfill, students’ confidence in their time estimates far exceeded their accomplishments.”18 This work has an interesting twist. While people are notoriously poor at guessing when they’ll finish their own projects, they’re pretty good at guessing about other people. In fact, the planning fallacy embodies a broader principle. When people are forced to look at similar situations and see the frequency of success, they tend to predict more accurately. If you want to know how something
FIGURE 1-2
FIGURE 1-2
人们以为自己完成任务的时间,与实际完成的时间之间,存在着巨大的差距。
There’s a huge gap between when people believe they will complete a task and when they actually do
Projected
Projected
Projected
Projected
Projected Actual
Projected Actual
Probability 50% 75% 99%
Probability 50% 75% 99%
Actual Actual
Actual Actual
13% 19% 45%
13% 19% 45%
Task completion
Task completion
来源:Roger Buehler、Dale Griffin 和 Michael Ross,《“关于时间”:工作与爱情中的乐观预测》,载于《欧洲社会心理学评论》第 6 卷,Wolfgang Stroebe 与 Miles Hewstone 编(英国奇切斯特:John Wiley & Sons,1995 年),第 1–32 页。
Source: Roger Buehler, Dale Griffin, and Michael Ross, “It’s About Time: Optimistic Predictions in Work and Love,” in European Review of Social Psychology, vol. 6, ed. Wolfgang Stroebe and Miles Hewstone (Chichester, UK: John Wiley & Sons, 1995), 1–32.
想知道自己会是什么结局,就去看看身处同样处境的人最终怎么样了。哈佛大学心理学家丹尼尔·吉尔伯特琢磨过,为什么人们不那么常借助外部视角来做判断:“既然这个简单的办法有如此令人印象深刻的力量,我们理应期待人们会不遗余力地去用它。但事实并非如此。”原因在于,大多数人都觉得自己与众不同、比身边的人更优秀。既然你现在知道了内部视角和外部视角如何影响人们的决策方式,你就会在随处可见的地方看到它。在商界,它会表现为对新品开发需要多久、并购交易成功的概率、股票组合跑赢市场的可能性抱有不合理的乐观。在个人生活中,你会看到父母坚信自己七岁的孩子注定能拿到大学体育奖学金,人们在争论电子游戏对孩子有什么影响,以及翻新一间厨房的时间和成本。
is going to turn out for you, look at how it turned out for others in the same situation. Daniel Gilbert, a psychologist at Harvard University, ponders why people don’t rely more on the outside view, “Given the impressive power of this simple technique, we should expect people to go out of their way to use it. But they don’t.” The reason is most people think of themselves as different, and better, than those around them.19 Now that you are aware of how the inside-outside view influ-ences the way people make decisions, you’ll see it everywhere. In the business world, it will show up as unwarranted optimism for how long it takes to develop a new product, the chance that a merger deal succeeds, and the likelihood a portfolio of stocks will do better than the market. In your personal life, you’ll see it in the parents who believe their seven-year-old is destined for a college sports scholarship, debates about what impact video games have on kids, and the time and cost it will take to remodel a kitchen.
就连那些本该更明白的人也常常忘记参考外部视角。多年前,丹尼尔·卡尼曼(Daniel Kahneman)组建了一个团队,编写一套高中判断与决策教学课程。这个团队里既有经验丰富的教师,也有缺乏经验的老师,还有教育学院的院长。大约一年后,他们才写了几章课本内容,并开发出了一些示范课程。
Even people who should know better forget to consult the outside view. Years ago, Daniel Kahneman assembled a group to write a curriculum to teach judgment and decision making to high school students. Kahneman’s group included a mix of experienced and inexperienced teachers as well as the dean of the school of education. After about a year, they had written a couple of chapters for the textbook and had developed some sample lessons.
在一次周五下午的研讨会上,这些教育工作者讨论了如何从团队中获取信息以及如何思考未来。他们知道,最好的办法是让每个人独立表达自己的观点,然后将这些观点汇总成共识。卡尼曼决定让这个练习变得具体——他请每位成员估算该团队向教育部提交教科书草稿的日期。卡尼曼发现,估算结果集中在两年左右,包括院长在内的所有人给出的时间都在 18 到 30 个月之间。这时卡尼曼突然想到,院长曾经参与过类似项目。当被问及时,院长说自己了解不少类似的团队,其中也包括曾经编写过教材的组。
During one of their Friday afternoon sessions, the educators discussed how to elicit information from groups and how to think about the future. They knew that the best way to do this was for each per-son to express his or her view independently and to combine the views into a consensus. Kahneman decided to make the exercise tangible by asking each member to estimate the date the group would deliver a draft of the textbook to the Ministry of Education. Kahneman found that the estimates clustered around two years and that everyone, including the dean, estimated between eighteen and thirty months. It then occurred to Kahneman that the dean had been involved in similar projects. When asked, the dean said he knew of a number of similar groups, including ones that had worked
在生物学和数学课程上。于是卡尼曼问了他那个显而易见的问题:“他们花了多长时间才完成?”
on the biology and mathematics curriculum. So Kahneman asked him the obvious question: “How long did it take them to finish?”
院长脸红了,随后回答说,在启动类似项目的团队中,有 40% 从未完成,而且所有团队完成时间都不少于七年。卡尼曼看到,要调和院长对这个团队的乐观回答与他所了解的其他团队的不足,只有一种可能,于是他问,这个团队和其他团队相比怎么样。停顿了一下,院长回答道:“低于平均水平,但差距不大。”20
The dean blushed and then answered that 40 percent of the groups that had started similar programs had never finished, and that none of the groups completed it in less than seven years. Seeing only one way to reconcile the dean’s optimistic answer about this group with his knowledge of the shortcomings of the other groups, Kahneman asked how good this group was compared with the others. After a pause, the dean responded, “Below average, but not by much.”20
如何将外部视角融入你的决策中
How to Incorporate the Outside View into Your Decisions
卡尼曼和与他长期合作的心理学家阿莫斯·特沃斯基曾发表过一个多步骤流程,帮助你运用外部视角。21 我将他们的五个步骤浓缩为四个,并加入了自己的一些思考。以下是这四个步骤:
Kahneman and Amos Tversky, a psychologist who had a long col-laboration with Kahneman, published a multistep process to help you use the outside view.21 I have distilled their five steps into four and have added some thoughts. Here are the four steps:
1. 选定一个参照系。找出一组情形或一个参照系,这个范围要足够大以保证统计显著性,又要足够窄以便对眼前的决策有分析价值。这一任务通常既是科学也是艺术,对于少有先例的问题尤其棘手。但对于常见的决策——即使对你而言不常见——确定一个参照系并不困难。要注意细节。以并购为例。我们知道在多数并购中,收购方股东是亏钱的。但仔细审视数据会发现,市场对现金交易、溢价幅度小的交易,反应要明显优于对高溢价、换股方式的交易。因此,公司如果了解哪些交易更容易成功,就能提高从并购中获利的概率。
1. Select a reference class. Find a group of situations, or a reference class, that is broad enough to be statistically signifi-cant but narrow enough to be useful in analyzing the decision that you face. The task is generally as much art as science, and is certainly trickier for problems that few people have dealt with before. But for decisions that are common—even if they are not common for you—identifying a reference class is straightforward. Mind the details. Take the example of mergers and acquisitions. We know that the shareholders of acquiring companies lose money in most mergers and acquisitions. But a closer look at the data reveals that the market responds more favorably to cash deals and those done at small premiums than to deals financed with stock at large premiums. So companies can improve their chances of making money from an acquisition by knowing what deals tend to succeed.
2. 评估结果的分布。一旦你有了一个参考类别,就仔细审视其中的成功率和失败率。
2. Assess the distribution of outcomes. Once you have a reference class, take a close look at the rate of success and fail-
当然。比如,Big Brown 处于当时的位置时,不到六分之一的马最终赢得了三冠王。研究分布情况,注意平均结果、最常见的结果,以及极端的成功或失败。
ure. For example, fewer than one of six horses in Big Brown’s position won the Triple Crown. Study the distribution and note the average outcome, the most common outcome, and extreme successes or failures.
在《满屋》一书中,哈佛大学古生物学家斯蒂芬·杰伊·古尔德展示了了解结果分布的重要性——起因是他的医生诊断出他患了间皮瘤。医生告诉他,被诊断出这种罕见癌症的人中,有一半只能活八个月(更准确地说,中位生存期是八个月),这听起来就像死刑宣判。但古尔德很快意识到,虽然一半患者在八个月内离世,但另一半患者却能活得更久很多。由于他确诊时相对年轻,他很有机会成为幸运者之一。古尔德写道:“我问对了问题,也找到了答案。在那种情况下,我几乎可以肯定获得了最宝贵的礼物——充裕的时间。”古尔德又活了二十年。22 另外两个问题值得一提。要使参考类别有效,成功与失败的统计概率必须随时间推移保持合理稳定。如果系统的性质发生变化,从过去数据中推断结论就可能产生误导。这在个人理财中是一个重要问题,因为理财顾问会基于历史统计数字为客户提出资产配置建议。由于市场的统计特性会随时间变化,投资者最终可能持有错误的资产组合。
In his book Full House, Stephen Jay Gould, who was a paleontologist at Harvard University, showed the impor-tance of knowing the distribution of outcomes after his doctor diagnosed him with mesothelioma. His doctor explained that half of the people diagnosed with the rare cancer lived only eight months (more technically, the median mortality was eight months), seemingly a death sentence. But Gould soon realized that while half the patients died within eight months, the other half went on to live much longer. Because of his relatively young age at diagnosis, there was a good chance he would be one of the fortunate ones. Gould wrote, “I had asked the right question and found the answers. I had obtained, in all probability, the most precious of all possible gifts in the circumstances—substantial time.” Gould lived another twenty years.22 Two other issues are worth mentioning. The statistical rate of success and failure must be reasonably stable over time for a reference class to be valid. If the properties of the system change, drawing inference from past data can be mis-leading. This is an important issue in personal finance, where advisers make asset allocation recommendations for their clients based on historical statistics. Because the statistical properties of markets shift over time, an investor can end up with the wrong mix of assets.
同时留意那些微小扰动就可能导致大规模变化的系统。在这些系统中,因果关系难以确定,因此借鉴过往经验也更为困难。像电影或书籍这类依赖爆款产品驱动的业务,就是很好的例子。制片
Also keep an eye out for systems where small perturba-tions can lead to large-scale change. Since cause and effect are difficult to pin down in these systems, drawing on past experiences is more difficult. Businesses driven by hit prod-ucts, like movies or books, are good examples. Producers
而出版商向来极难预测业绩,因为成功与失败很大程度上取决于社会影响力——这本身就是一个不可预测的现象。
and publishers have a notoriously difficult time anticipating results, because success and failure is based largely on social influence, an inherently unpredictable phenomenon.
3. 做出预测。掌握了参照类的数据,包括对结果分布的认识之后,你就有条件做出预测了。思路是估算自己成功和失败的概率。由于我讨论过的所有原因,你的预测很可能过于乐观。
3. Make a prediction. With the data from your reference class in hand, including an awareness of the distribution of outcomes, you are in a position to make a forecast. The idea is to estimate your chances of success and failure. For all the reasons that I’ve discussed, the chances are good that your prediction will be too optimistic.
有时当你找到了合适的参照类别,你会发现成功率并不高。因此,为了提高成功几率,你必须做出与其他人不同的选择。一个例子是美国国家橄榄球联盟(NFL)教练在比赛关键时刻的战术决策,包括四档进攻、开球以及两分转换尝试。与许多其他运动类似,处理这些情况的传统方式是一代又一代教练传下来的。但这种陈旧的决策过程意味着得分更少、赢下的比赛也更少。
Sometimes when you find the right reference class, you see the success rate is not very high. So to improve your chance of success, you have to do something different than everyone else. One example is the play calling of National Football League coaches in critical game situations including fourth downs, kickoffs, and two-point conversion attempts. As in many other sports, conventional ways to decide about these situations are handed down from one generation of coaches to the next. But this stale decision-making process means scoring fewer points and winning fewer games.
印第安纳大学的天体物理学家查克·鲍尔与前世界双陆棋冠军弗兰克·弗里戈共同开发了一款名为“宙斯”的计算机程序,用以评估职业橄榄球教练的战术决策。宙斯采用了在双陆棋和象棋程序中获得成功的相同建模技术,其开发者将统计数据及教练的行为特征加载到程序中。鲍尔和弗里戈发现,在拥有三十二支球队的联盟中,仅有四支球队的关键决策在超过一半的情况下与宙斯的判断一致,而有九支球队的决策与宙斯的吻合度不足四分之一。据宙斯估算,这些糟糕的决策可能使一支球队每年损失超过一场胜利,在十六场比赛的赛季中,这一代价相当沉重。
Chuck Bower, an astrophysicist at Indiana University, and Frank Frigo, a former world backgammon champion, created a computer program called Zeus to assess the play-calling decisions of pro football coaches. Zeus uses the same modeling techniques that have succeeded in backgammon and chess programs, and the creators loaded it with statistics and the behavioral traits of coaches. Bower and Frigo found that only four teams in the thirty-two-team league made crucial decisions that agreed with Zeus over one-half of the time, and that nine teams made decisions that con-curred less than one-quarter of the time. Zeus estimates that these poor decisions can cost a team more than one victory per year, a large toll in a sixteen-game season.
多数教练固守传统智慧,因为那是他们学到的,而且他们对此习以为常——
Most coaches stick to the conventional wisdom, because that is what they have learned and they are averse to the
脱离过去做法的负面后果。但宙斯(Zeus)表明,外部视角可以为愿意打破传统的教练带来更多胜利。
perceived negative consequences of breaking from past practice. But Zeus shows that the outside view can lead to more wins for the coach willing to break with tradition.
这是一个愿意三思而后行的教练所面临的机遇。
This is an opportunity for coaches willing to think twice.23
4. 评估预测的可靠性并微调。我们决策能力的高低,很大程度上取决于预测的对象。例如,天气预报员在预测明天温度方面做得相当出色。而图书出版商则不擅长挑出畅销书,除了少数几位畅销书作家的作品之外。成功预测的记录越差,你就越应该将你的预测向均值(或其他相关统计指标)靠拢。当因果关系清晰时,你可以对自己的预测更有信心。
4. Assess the reliability of your prediction and fine-tune. How good we are at making decisions depends a great deal on what we are trying to predict. Weather forecasters, for instance, do a pretty good job of predicting what the temperature will be tomorrow. Book publishers, on the other hand, are poor at picking winners, with the exception of those books from a handful of best-selling authors. The worse the record of successful prediction is, the more you should adjust your prediction toward the mean (or other relevant statistical measure). When cause and effect is clear, you can have more confidence in your forecast.
内部-外部视角的主要教训是:虽然决策者倾向于关注独特性,但最好的决策往往源于共性。别误会我的意思。我并不是在提倡平淡、缺乏想象力、模仿或规避风险的决策。我是说,基于与日常情况相似的场景,存在着大量有用的信息。忽视这些信息,对我们自身是不利的。关注这些宝贵的信息资源,能帮助你做出更有效的决策。下次再有“三冠王”候选者以极其乐观的赔率出赛时,请记住这番讨论。
The main lesson from the inside-outside view is that while decision makers tend to dwell on uniqueness, the best decisions often derive from sameness. Don’t get me wrong. I’m not advocating for bland, unimaginative, imitative, or risk-free decisions. I am saying there is a wealth of useful information based on situations that are similar to the ones that we face every day. We ignore that information to our own detriment. Paying attention to that wealth of information will help you make more effective decisions. Remember this discussion the next time a contender for the Triple Crown goes off at highly optimistic odds.