Leadership retreats and decision-making
Why can group decisions outperform individual decisions?
For Managers, facilitators, leadership teams, educators, and learning designers
- Last updated
- Reading goal
- Understand the conditions behind better group judgment and interpret individual-versus-team results responsibly
Planning answer
A group decision can outperform individual decisions when people begin with sufficiently independent judgments, bring different useful information, explain their reasoning, and combine contributions in a way that rewards accuracy rather than status. It is not automatic. Social influence can erase useful differences, dominant voices can crowd out evidence, and teams often discuss information everyone already knows instead of facts held by only one member.
Decisions to settle
- Always name the comparison: average, median, or best individual.
- Independent input protects information that early consensus might erase.
- Groups improve when unique evidence reaches the discussion and can be evaluated.
- A narrower spread is more consistent, but not necessarily more accurate.
- A before-and-after score describes change; it does not prove what caused it.
Under what conditions can a group make a better decision?
Groups have an advantage when members hold partly different information or approaches and the task lets them demonstrate why an answer is better. The group must actually surface those contributions. Independent first judgments preserve starting differences; explanation makes a proposal inspectable; and a clear decision rule helps the team revise for reasons rather than simply follow confidence or seniority.
- Members form an initial view before seeing the emerging consensus.
- The group has relevant diversity of information, experience, or problem-solving approach.
- People can explain and test the reasoning behind a choice.
- Unique information is actively requested instead of waiting for someone to volunteer it.
- The group distinguishes expertise and evidence from confidence, title, or airtime.
- The final method combines contributions rather than defaulting to the first suggestion.
In a 2006 experiment on highly demonstrable letters-to-numbers problems, groups of three, four, and five solved the tasks in fewer trials than the best of equivalent numbers of individuals. That is strong evidence for those structured, solvable problems—not proof that groups beat their best member on ambiguous workplace judgments.Journal of Personality and Social Psychology via PubMed
Is discussion the same as the wisdom of crowds?
No. Classic crowd wisdom normally aggregates judgments that retain some independence. Face-to-face discussion creates social influence: people exchange reasons, but they also react to one another's confidence and preferences. Discussion can improve the inputs, reduce the independence between them, or do both. The design of the process matters as much as the number of people.
A 2018 live-crowd study asked 5,180 people general-knowledge questions, placed them in groups of five for deliberation, and then aggregated several group consensus estimates. In that design, aggregating a small number of group consensuses was more accurate than aggregating the original crowd estimates. The result depends on a two-level method—small-group debate followed by aggregation—and should not be reduced to “any meeting improves accuracy.”Nature Human Behaviour
| Method | What it preserves | Main risk |
|---|---|---|
| Average independent estimates | Independence and error cancellation | A simple mean can be distorted by scale, bias, or extreme values |
| Discuss and reach one consensus | Reason exchange and coordinated choice | Influence can remove useful disagreement or amplify a persuasive error |
| Discuss in small groups, then aggregate | Local reasoning plus diversity between groups | Benefits depend on independent groups and a suitable aggregation rule |
| Select an expert or best member | A strong identifiable judgment | The group may misidentify expertise or ignore complementary information |
Swipe or use the left and right arrow keys to see more
Why can a group decision be worse?
Discussion can create correlated errors. In an experiment with 144 participants performing simple estimation tasks, even mild exposure to other people's estimates narrowed the range of answers without improving collective error. The crowd became more consistent in the everyday sense of giving similar answers, but not more accurate.Proceedings of the National Academy of Sciences via PubMed
Groups also favor shared information. A meta-analysis covering 65 hidden-profile studies, 101 independent effects, and 3,189 groups found that groups mentioned substantially more common than unique information. Hidden-profile groups were eight times less likely to find the solution than groups whose members all had the full information; measures of pooling unique information were positively related to decision quality. These studies use a particular experimental paradigm, but they identify a practical facilitation risk: a fact known by one person may never enter the team answer.Personality and Social Psychology Review via PubMed
- The first proposal anchors the conversation.
- A senior or confident member is treated as correct without testing the reason.
- People repeat common facts because those facts receive quick agreement.
- Unique evidence arrives after the group has publicly committed to a choice.
- Conflict avoidance turns consensus into the objective instead of decision quality.
- Everyone relies on the same source, so errors are correlated rather than canceled.
What does it mean to say the group “did better”?
The claim is incomplete until it names a benchmark. A team can beat the average individual while still finishing behind its best member. It can produce a higher average result with greater variability, or become more consistent around a worse answer. Each comparison answers a different question.
| Claim | Required comparison | What it does not establish |
|---|---|---|
| The group beat the average individual | Group score versus the arithmetic mean of comparable individual scores | That the group beat the median or best person |
| The group beat the median individual | Group score versus the middle comparable individual score | That it beat the average or best person |
| The group beat the best individual | Group score versus the strongest comparable individual score | That discussion will repeat the result on another task |
| The group was more consistent | A smaller spread, variance, or standard deviation under the same measure | That the center of the scores was better |
| Discussion caused improvement | A design that can rule out plausible alternative causes | What a simple before-and-after difference can show |
Swipe or use the left and right arrow keys to see more
Lower variance is not automatically desirable. It means observations are more tightly clustered. Whether that is useful depends on where they are clustered, whether lower or higher scores are better, what target matters, and how costly extreme errors would be. Likewise, “highest potential” is too vague for a result statement; name the best observed score, the theoretical limit, or the benchmark you actually measured.
How can a facilitator improve the decision process?
Collect independent first views
Ask each person to record a choice and reason before open discussion. Do not display the emerging distribution while people are still forming that view.
Invite unique information
Ask what one person knows that the rest of the group may not. Round-robin evidence before debating preferred answers.
Separate reasons from people
Write the evidence and tradeoffs where the group can inspect them without tying every point to status or personality.
Test the leading answer
Ask what would make the current choice wrong, which assumption is carrying the most weight, and what evidence has not been reconciled.
Record the comparison honestly
Use the same task, reference, and scoring direction for individual and group answers. Report missing work as unavailable rather than zero.
Debrief the process
Ask which information changed the decision, who influenced the revision, what was overlooked, and whether the team would use the same process again.
What the comparison can—and cannot—show
The current Decision Challenge records completed individual rankings before the team submits one shared ranking. It compares the team with the mean distance of its own completed members, so “closer than the average member” has one precise meaning.
The result can be closer, unchanged, farther, or unavailable. It cannot show what caused the change, prove the team beat its best member, or diagnose a lasting team trait. The scoring page has the formula and interpretation limits.
Related practical questions
- Does a team usually beat its best member?
- Not necessarily. Evidence that groups beat their best member comes from specific tasks and conditions. On ambiguous decisions, the best member may outperform consensus, and the group may not know who that person is.
- Is a unanimous decision a high-quality decision?
- Unanimity shows agreement, not accuracy. Ask which evidence was considered, whether unique information surfaced, and whether people formed views before social influence.
- Does less variance mean a better outcome?
- No. Less variance means a tighter spread. You must also inspect the average or target, the score direction, and the cost of extreme outcomes.
- Can a before-and-after team exercise show causation?
- A simple comparison can show change under the exercise's scoring rule. It cannot isolate discussion as the cause without a stronger study design.
Sources and limits
The cited studies cover specific tasks, group sizes, and information conditions; they do not prove a general group advantage or an Escape Scenario effect. We rechecked the Decision Challenge formula and limits on August 29, 2026.
- Groups Perform Better Than the Best Individuals on Letters-to-Numbers Problems: Effects of Group SizeJournal of Personality and Social Psychology via PubMed, 2006. An experiment on highly demonstrable letters-to-numbers problems; the task and group-size conditions limit generalization.
- How Escape Scenario Scoring WorksEscape Scenario. The public explanation of the current in-room flow, Decision Challenge distance formula, comparison baseline, and interpretation limits.