Consulting interviewers at McKinsey, BCG, Bain and adjacent firms grade candidates against structured evaluation criteria, not a vague impression. The firms do not publish their rubrics, and there is no universal MBB scorecard. What follows is the CasingLab six-dimension rubric: our own assessment framework, built by ex-MBB consultants to capture the skills interviewer and candidate accounts most commonly describe being evaluated. This guide breaks down those six dimensions, the evidence that earns top marks, the habits that trigger low scores, and a self-assessment tool to benchmark your own performance.
Why the Scorecard Matters
Firms standardise their evaluation for two reasons: consistency across interviewers and defensibility of hiring decisions. An interviewer is expected to point to observed behaviour, not "bad vibes". That is good news for candidates, because it means every rejection has a reason you can train against.
Understanding the rubric also compresses prep time. Instead of diffusely "getting better at cases", you can target the one or two dimensions where your current score is pulling the average down.
The Six Dimensions and Weights
No firm publishes its scorecard, and the exact format varies by firm and office. The table below is the CasingLab scoring rubric, built by our ex-MBB team to reflect the dimensions interviewers consistently describe evaluating and how heavily each one tends to count.
| Dimension | CasingLab Weight | What It Measures |
|---|---|---|
| Structure | 20% | MECE decomposition, hypothesis-driven thinking |
| Quantitative | 20% | Mental math, data interpretation, estimation |
| Business judgment | 20% | Prioritisation, commercial awareness, pattern recognition |
| Communication | 15% | Top-down delivery, conciseness, signposting |
| Creativity | 15% | Brainstorming diversity, non-obvious ideas |
| Synthesis | 10% | CEO-ready recommendation, conviction |
CasingLab scores each dimension 1 to 10; candidate accounts describe firms using similar scales, often 1 to 5. On any version, a single dimension near the bottom of the scale typically sinks the overall evaluation regardless of how the others look. For the band-by-band anchors (what a 1, a 3 and a 5 look like on each dimension) and a printable version, see the full consulting interview scorecard reference.
Structure (20%)
Structure is the foundation, assessed in the first 60-90 seconds. Top scores go to candidates who build a MECE, case-specific tree with prioritisation and an explicit hypothesis. For the method behind strong structures, see our 60-second structure guide.
What a top answer looks like
- Restates the core question in one sentence
- 3-4 MECE buckets with industry-specific sub-questions
- Prioritises one branch and justifies why
- States a clear, falsifiable hypothesis
What a weak answer looks like
- Recites a memorised framework (Porter, 4Ps, 3Cs)
- Lists seven buckets with heavy overlap
- Dives into analysis before naming all buckets
- Uses generic labels like "internal and external factors"
Quantitative (20%)
Quantitative scoring is less about arithmetic speed and more about setup and interpretation. Interviewers want to see you write a clean equation, compute under narration, and then interpret what the number means.
High scores come from candidates who state rounding choices out loud, estimate before calculating, and translate results into business implications. For worked estimation practice, see our market sizing examples.
Business Judgment (20%)
Business judgment is the dimension partners care most about, because it maps directly to client value. It shows up in how you prioritise, which data you request, and whether your recommendations pass a commercial smell test.
Signals of Strong Judgment
- Chooses the highest-leverage bucket to investigate first
- Asks for data that could change the recommendation
- Challenges assumptions without derailing the case
- Translates numbers into operational implications
Communication (15%)
Communication is scored continuously throughout the case. Top performers are top-down, concise and signposted. They say "I see three points; first... second... third" and pause for acknowledgment.
Hedging language ("I think maybe", "perhaps we could") drags scores down fast. So does monologuing beyond 60 seconds without checking in.
Creativity (15%)
Creativity typically surfaces in brainstorming prompts ("give me ideas for how the client could grow"). Scores track both quantity and diversity.
| Band | Typical Output |
|---|---|
| Top | 10+ ideas across 3-4 distinct categories |
| Strong | 7-8 ideas across 2-3 categories |
| Average | 5-6 mostly obvious ideas |
| Weak | 3-4 ideas, all from the same angle |
A structured "organic, inorganic and adjacencies" scaffold beats unstructured list-dumping every time.
Synthesis (10%)
Synthesis is the close: 60 seconds, recommendation-first, three reasons with numbers, one risk and a next step. A weak synthesis after a strong case still lowers the overall score significantly. For the exact formula, see our case interview synthesis guide.
The Question Behind the Scores
Beneath the dimension scores, interviewers consistently describe asking themselves one deciding question: would I be comfortable putting this person in front of my client next week?
A yes typically becomes an offer. A no typically becomes a rejection, even if the dimension scores look decent on paper. That judgment is the interviewer's override, and it is the reason a single glaring weakness matters more than three strong dimensions.
Insider Tip: Candidates obsess about perfecting their strong dimensions. The higher-leverage move is lifting the weakest dimension out of the danger zone.
How to Self-Assess After a Practice Case
Use this rubric after every practice session. Be honest; the scorecard is only useful if it reflects how an interviewer would have graded you. A printable self-assessment checklist with the observable behaviours for each dimension is available as a standalone reference, no account needed.
- Structure: Was it MECE, specific, prioritised and hypothesis-led?
- Quant: Did I set up cleanly, narrate, and interpret the result?
- Judgment: Did I chase the right data first?
- Communication: Was I top-down and under 60 seconds per answer?
- Creativity: Did I generate 8+ ideas across 3+ categories when asked?
- Synthesis: Did I open with the recommendation and quantify each reason?
Scoring below 3 on any dimension flags your next week's prep focus. If self-scoring feels unreliable, you can see your own six-dimension scorecard after each AI case and compare it against your honest self-assessment.
Common Mistakes to Avoid
- Self-scoring too generously (peers and AI scoring almost always rate lower)
- Practising your strengths because they feel good
- Ignoring the weakest dimension because it feels unfixable
- Treating the final synthesis as optional
- Forgetting that judgment is scored throughout, not just at the end
FAQ
Do all firms use the same scorecard?
No firm publishes its weightings, so any firm-by-firm comparison is inference from interview styles rather than documented fact. What the firms do publish about their interview formats differs: see our McKinsey interview guide and Bain interview guide for what each firm actually says it assesses.
Can one strong dimension compensate for a weak one?
Rarely, by interviewer accounts: a very low score on any dimension tends to weigh heavily regardless of strong scores elsewhere, because it fails the comfortable-in-front-of-a-client test. No firm documents a formal rule here.
How many practice cases should I do per week?
Three to five full cases per week, each followed by a rubric self-assessment, is the cadence we recommend over six to eight weeks. Our 8-week case interview preparation plan sequences that cadence week by week.
Should I focus on fixing weaknesses or sharpening strengths?
Fix weaknesses first until every dimension is a 3 or higher, then sharpen. The marginal gain from 4 to 5 on a strong dimension is smaller than from 2 to 3 on a weak one.
How do I get reliable feedback without a live interviewer?
Use AI-scored mock interviews that grade all six dimensions with rationale and transcript references. It is the fastest way to convert practice into measurable improvement.
Sources & Further Reading
- McKinsey & Company, Our Leadership: official commentary on partnership and hiring philosophy.
- McKinsey & Company, Featured Insights: background on the firm's approach to talent and capability building.
- Bain & Company, Insights: material on judgment-led problem solving.
- Harvard Business Review, Decision Making and Problem Solving topic archive: applied research on structured judgment and assessment.
Get a six-dimension scorecard after every AI-powered practice session. CasingLab is now open. Start the free diagnostic and see your readiness score in five minutes.