Consulting Interview Scorecard: The 6 Dimensions Explained

The consulting interview scorecard is the hidden grading rubric that decides every offer at McKinsey, BCG, Bain and adjacent firms. Interviewers are not forming a vague impression; they are scoring you on six specific dimensions on a defined scale and writing one final sentence that determines the outcome. This guide breaks down each dimension, the evidence that earns top marks, the habits that trigger low scores, and a self-assessment tool to benchmark your own performance.

Why the Scorecard Matters

Firms standardised scorecards for two reasons: consistency across interviewers and defensibility of hiring decisions. A partner cannot ding a candidate for "bad vibes"; they must point to observed behaviour mapped to a rubric. That is good news for candidates, because it means every rejection has a reason you can train against.

Understanding the rubric also compresses prep time. Instead of diffusely "getting better at cases", you can target the one or two dimensions where your current score is pulling the average down.

The Six Dimensions and Weights

Scorecards vary slightly by firm, but the six dimensions and their approximate weights are remarkably consistent.

Dimension Typical Weight What It Measures
Structure 20% MECE decomposition, hypothesis-driven thinking
Quantitative 20% Mental math, data interpretation, estimation
Business judgment 20% Prioritisation, commercial awareness, pattern recognition
Communication 15% Top-down delivery, conciseness, signposting
Creativity 15% Brainstorming diversity, non-obvious ideas
Synthesis 10% CEO-ready recommendation, conviction

The dimensions are scored 1 to 10 (some firms use 1 to 5). A single dimension at 2 or below typically sinks the overall score regardless of how the others look.

Structure (20%)

Structure is the foundation, assessed in the first 60-90 seconds. Top scores go to candidates who build a MECE, case-specific tree with prioritisation and an explicit hypothesis. For the method behind strong structures, see our 60-second structure guide.

Evidence of a 5/5

  • Restates the core question in one sentence
  • 3-4 MECE buckets with industry-specific sub-questions
  • Prioritises one branch and justifies why
  • States a clear, falsifiable hypothesis

Evidence of a 2/5

  • Recites a memorised framework (Porter, 4Ps, 3Cs)
  • Lists seven buckets with heavy overlap
  • Dives into analysis before naming all buckets
  • Uses generic labels like "internal and external factors"

Quantitative (20%)

Quantitative scoring is less about arithmetic speed and more about setup and interpretation. Interviewers want to see you write a clean equation, compute under narration, and then interpret what the number means.

High scores come from candidates who state rounding choices out loud, estimate before calculating, and translate results into business implications. For worked estimation practice, see our market sizing examples.

Business Judgment (20%)

Business judgment is the dimension partners care most about, because it maps directly to client value. It shows up in how you prioritise, which data you request, and whether your recommendations pass a commercial smell test.

Signals of Strong Judgment

  • Chooses the highest-leverage bucket to investigate first
  • Asks for data that could change the recommendation
  • Challenges assumptions without derailing the case
  • Translates numbers into operational implications

Communication (15%)

Communication is scored continuously throughout the case. Top performers are top-down, concise and signposted. They say "I see three points; first... second... third" and pause for acknowledgment.

Hedging language ("I think maybe", "perhaps we could") drags scores down fast. So does monologuing beyond 60 seconds without checking in.

Creativity (15%)

Creativity typically surfaces in brainstorming prompts ("give me ideas for how the client could grow"). Scores track both quantity and diversity.

Score Typical Output
5/5 10+ ideas across 3-4 distinct categories
4/5 7-8 ideas across 2-3 categories
3/5 5-6 mostly obvious ideas
2/5 3-4 ideas, all from the same angle

A structured "organic, inorganic and adjacencies" scaffold beats unstructured list-dumping every time.

Synthesis (10%)

Synthesis is the close: 60 seconds, recommendation-first, three reasons with numbers, one risk and a next step. A weak synthesis after a strong case still lowers the overall score significantly. For the exact formula, see our case interview synthesis guide.

The Final Line on Every Scorecard

At the bottom of every scorecard sits a single question: "Would I be comfortable putting this person in front of my client next week?"

Yes typically becomes an offer. No typically becomes a rejection, even if the dimension scores look decent on paper. The line is the partner's override and it is the reason a single glaring weakness matters more than three strong dimensions.

Insider Tip: Candidates obsess about reaching 5/5 on their strong dimensions. The higher-leverage move is lifting the weakest dimension from 2 to 3.

How to Self-Assess After a Practice Case

Use this rubric after every practice session. Be honest; the scorecard is only useful if it reflects how an interviewer would have graded you.

  • Structure: Was it MECE, specific, prioritised and hypothesis-led?
  • Quant: Did I set up cleanly, narrate, and interpret the result?
  • Judgment: Did I chase the right data first?
  • Communication: Was I top-down and under 60 seconds per answer?
  • Creativity: Did I generate 8+ ideas across 3+ categories when asked?
  • Synthesis: Did I open with the recommendation and quantify each reason?

Scoring below 3 on any dimension flags your next week's prep focus. If self-scoring feels unreliable, you can see your own six-dimension scorecard after each AI case and compare it against your honest self-assessment.

Common Mistakes to Avoid

  • Self-scoring too generously (peers and AI scoring almost always rate lower)
  • Practising your strengths because they feel good
  • Ignoring the weakest dimension because it feels unfixable
  • Treating the final synthesis as optional
  • Forgetting that judgment is scored throughout, not just at the end

FAQ

Do all firms use the same scorecard?

The dimensions are similar but weightings shift. McKinsey emphasises structure and quant; BCG often rewards creativity more; Bain tends to weight judgment and synthesis heavily. See our McKinsey interview guide and Bain interview guide for firm-specific tilts.

Can one strong dimension compensate for a weak one?

Rarely. A 2 on any dimension typically triggers a no-offer regardless of 5s elsewhere, because it fails the "comfortable in front of a client" test.

How many practice cases should I do per week?

Three to five full cases per week, each followed by a rubric self-assessment, is the typical offer-winning cadence over six to eight weeks. Our 8-week case interview preparation plan sequences that cadence week by week.

Should I focus on fixing weaknesses or sharpening strengths?

Fix weaknesses first until every dimension is a 3 or higher, then sharpen. The marginal gain from 4 to 5 on a strong dimension is smaller than from 2 to 3 on a weak one.

How do I get reliable feedback without a live interviewer?

Use AI-scored mock interviews that grade all six dimensions with rationale and transcript references. It is the fastest way to convert practice into measurable improvement.

Sources & Further Reading

  1. McKinsey & Company, Our Leadership: official commentary on partnership and hiring philosophy.
  2. McKinsey & Company, Featured Insights: background on the firm's approach to talent and capability building.
  3. Bain & Company, Insights: material on judgment-led problem solving.
  4. Harvard Business Review, Decision Making and Problem Solving topic archive: applied research on structured judgment and assessment.

Get a six-dimension scorecard after every AI-powered practice session. CasingLab is now open. Start the free diagnostic and see your readiness score in five minutes.