Skip to content

Leadership retreats and decision-making

Why can group decisions outperform individual decisions?

For Managers, facilitators, leadership teams, educators, and learning designers

Last updated
Reading goal
Understand the conditions behind better group judgment and interpret individual-versus-team results responsibly

Planning answer

A group decision can outperform individual decisions when people begin with sufficiently independent judgments, bring different useful information, explain their reasoning, and combine contributions in a way that rewards accuracy rather than status. It is not automatic. Social influence can erase useful differences, dominant voices can crowd out evidence, and teams often discuss information everyone already knows instead of facts held by only one member.

Decisions to settle

  • Always name the comparison: average, median, or best individual.
  • Independent input protects information that early consensus might erase.
  • Groups improve when unique evidence reaches the discussion and can be evaluated.
  • A narrower spread is more consistent, but not necessarily more accurate.
  • A before-and-after score describes change; it does not prove what caused it.

Under what conditions can a group make a better decision?

Groups have an advantage when members hold partly different information or approaches and the task lets them demonstrate why an answer is better. The group must actually surface those contributions. Independent first judgments preserve starting differences; explanation makes a proposal inspectable; and a clear decision rule helps the team revise for reasons rather than simply follow confidence or seniority.

  • Members form an initial view before seeing the emerging consensus.
  • The group has relevant diversity of information, experience, or problem-solving approach.
  • People can explain and test the reasoning behind a choice.
  • Unique information is actively requested instead of waiting for someone to volunteer it.
  • The group distinguishes expertise and evidence from confidence, title, or airtime.
  • The final method combines contributions rather than defaulting to the first suggestion.

In a 2006 experiment on highly demonstrable letters-to-numbers problems, groups of three, four, and five solved the tasks in fewer trials than the best of equivalent numbers of individuals. That is strong evidence for those structured, solvable problems—not proof that groups beat their best member on ambiguous workplace judgments.Journal of Personality and Social Psychology via PubMed

Is discussion the same as the wisdom of crowds?

No. Classic crowd wisdom normally aggregates judgments that retain some independence. Face-to-face discussion creates social influence: people exchange reasons, but they also react to one another's confidence and preferences. Discussion can improve the inputs, reduce the independence between them, or do both. The design of the process matters as much as the number of people.

A 2018 live-crowd study asked 5,180 people general-knowledge questions, placed them in groups of five for deliberation, and then aggregated several group consensus estimates. In that design, aggregating a small number of group consensuses was more accurate than aggregating the original crowd estimates. The result depends on a two-level method—small-group debate followed by aggregation—and should not be reduced to “any meeting improves accuracy.”Nature Human Behaviour

Different ways to produce a collective answer
MethodWhat it preservesMain risk
Average independent estimatesIndependence and error cancellationA simple mean can be distorted by scale, bias, or extreme values
Discuss and reach one consensusReason exchange and coordinated choiceInfluence can remove useful disagreement or amplify a persuasive error
Discuss in small groups, then aggregateLocal reasoning plus diversity between groupsBenefits depend on independent groups and a suitable aggregation rule
Select an expert or best memberA strong identifiable judgmentThe group may misidentify expertise or ignore complementary information

Why can a group decision be worse?

Discussion can create correlated errors. In an experiment with 144 participants performing simple estimation tasks, even mild exposure to other people's estimates narrowed the range of answers without improving collective error. The crowd became more consistent in the everyday sense of giving similar answers, but not more accurate.Proceedings of the National Academy of Sciences via PubMed

Groups also favor shared information. A meta-analysis covering 65 hidden-profile studies, 101 independent effects, and 3,189 groups found that groups mentioned substantially more common than unique information. Hidden-profile groups were eight times less likely to find the solution than groups whose members all had the full information; measures of pooling unique information were positively related to decision quality. These studies use a particular experimental paradigm, but they identify a practical facilitation risk: a fact known by one person may never enter the team answer.Personality and Social Psychology Review via PubMed

  • The first proposal anchors the conversation.
  • A senior or confident member is treated as correct without testing the reason.
  • People repeat common facts because those facts receive quick agreement.
  • Unique evidence arrives after the group has publicly committed to a choice.
  • Conflict avoidance turns consensus into the objective instead of decision quality.
  • Everyone relies on the same source, so errors are correlated rather than canceled.

What does it mean to say the group “did better”?

The claim is incomplete until it names a benchmark. A team can beat the average individual while still finishing behind its best member. It can produce a higher average result with greater variability, or become more consistent around a worse answer. Each comparison answers a different question.

Claims that should not be treated as interchangeable
ClaimRequired comparisonWhat it does not establish
The group beat the average individualGroup score versus the arithmetic mean of comparable individual scoresThat the group beat the median or best person
The group beat the median individualGroup score versus the middle comparable individual scoreThat it beat the average or best person
The group beat the best individualGroup score versus the strongest comparable individual scoreThat discussion will repeat the result on another task
The group was more consistentA smaller spread, variance, or standard deviation under the same measureThat the center of the scores was better
Discussion caused improvementA design that can rule out plausible alternative causesWhat a simple before-and-after difference can show

Lower variance is not automatically desirable. It means observations are more tightly clustered. Whether that is useful depends on where they are clustered, whether lower or higher scores are better, what target matters, and how costly extreme errors would be. Likewise, “highest potential” is too vague for a result statement; name the best observed score, the theoretical limit, or the benchmark you actually measured.

How can a facilitator improve the decision process?

  1. Collect independent first views

    Ask each person to record a choice and reason before open discussion. Do not display the emerging distribution while people are still forming that view.

  2. Invite unique information

    Ask what one person knows that the rest of the group may not. Round-robin evidence before debating preferred answers.

  3. Separate reasons from people

    Write the evidence and tradeoffs where the group can inspect them without tying every point to status or personality.

  4. Test the leading answer

    Ask what would make the current choice wrong, which assumption is carrying the most weight, and what evidence has not been reconciled.

  5. Record the comparison honestly

    Use the same task, reference, and scoring direction for individual and group answers. Report missing work as unavailable rather than zero.

  6. Debrief the process

    Ask which information changed the decision, who influenced the revision, what was overlooked, and whether the team would use the same process again.

What the comparison can—and cannot—show

The current Decision Challenge records completed individual rankings before the team submits one shared ranking. It compares the team with the mean distance of its own completed members, so “closer than the average member” has one precise meaning.

The result can be closer, unchanged, farther, or unavailable. It cannot show what caused the change, prove the team beat its best member, or diagnose a lasting team trait. The scoring page has the formula and interpretation limits.

See the in-room flow and scoring formulaLearn how Decision Challenge distance, completed-member baselines, ties, and discussion change are calculated.
See the in-room flow and scoring formula

Related practical questions

Does a team usually beat its best member?
Not necessarily. Evidence that groups beat their best member comes from specific tasks and conditions. On ambiguous decisions, the best member may outperform consensus, and the group may not know who that person is.
Is a unanimous decision a high-quality decision?
Unanimity shows agreement, not accuracy. Ask which evidence was considered, whether unique information surfaced, and whether people formed views before social influence.
Does less variance mean a better outcome?
No. Less variance means a tighter spread. You must also inspect the average or target, the score direction, and the cost of extreme outcomes.
Can a before-and-after team exercise show causation?
A simple comparison can show change under the exercise's scoring rule. It cannot isolate discussion as the cause without a stronger study design.

Sources and limits

The cited studies cover specific tasks, group sizes, and information conditions; they do not prove a general group advantage or an Escape Scenario effect. We rechecked the Decision Challenge formula and limits on August 29, 2026.

  1. Groups Perform Better Than the Best Individuals on Letters-to-Numbers Problems: Effects of Group SizeJournal of Personality and Social Psychology via PubMed, 2006. An experiment on highly demonstrable letters-to-numbers problems; the task and group-size conditions limit generalization.
  2. Aggregated Knowledge from a Small Number of Debates Outperforms the Wisdom of Large CrowdsNature Human Behaviour, 2018. A live-crowd study with 5,180 participants using individual general-knowledge estimates, five-person deliberation, and aggregation of group consensus estimates.
  3. How Social Influence Can Undermine the Wisdom of Crowd EffectProceedings of the National Academy of Sciences via PubMed, 2011. An experiment with 144 participants and simple estimation tasks showing that convergence can reduce diversity without improving collective error.
  4. Twenty-Five Years of Hidden Profiles in Group Decision Making: A Meta-AnalysisPersonality and Social Psychology Review via PubMed, 2012. A meta-analysis of 65 studies, 101 independent effects, and 3,189 groups examining common versus unique information and decision quality.
  5. How Escape Scenario Scoring WorksEscape Scenario. The public explanation of the current in-room flow, Decision Challenge distance formula, comparison baseline, and interpretation limits.