White-label quarterly planning for coaches · Live on your domain in about a weekPlanning ScorecardPricingFAQBook a demo
Method comparison

Prioritization methods compared: pairwise, weighted scoring, ICE and gut feel

Four ways to rank a quarter’s candidates, what each one is actually good at, and which one clients will finish without being chased.

8 min read · All guides

Key takeaways
  • Gut feel is fast and fine for the top two. It fails at the boundary — the difference between number four and number seven, which is where capacity runs out.
  • Weighted scoring is defensible on paper and routinely abandoned in practice because the arithmetic is tedious and the scale drifts.
  • ICE is quick and popular, but three subjective scores multiplied together hide a lot of guesswork.
  • Pairwise comparison is the only one most clients reliably finish, and it produces an order they will defend.

Why the method matters at all

It is tempting to treat ranking as a formality — everyone knows what matters most, so why formalise it? Because the top of the list is not where quarters are lost. They are lost at the boundary: the point where the fifth priority is accepted and the sixth is not, which is exactly the region where intuition is weakest and politics strongest.

A ranking method earns its keep at that boundary, and it earns it again three weeks later when someone asks why their initiative did not make the cut.

Gut feel

Fast, free, and genuinely reliable for the top one or two items. Experienced owners usually do know what the biggest thing is, and making them prove it with arithmetic can be condescending.

It fails in three predictable ways: recency bias, whoever spoke most recently, and status — the item belonging to the most senior person in the room drifts upward. It also leaves nothing behind. When the ranking is challenged later, there is no reasoning to point at, only a memory of a conversation.

Weighted scoring matrices

The spreadsheet approach: list candidates as rows, criteria as columns, score each cell, multiply by weights, total. Rigorous in principle, and the criteria discussion alone is often valuable.

In practice, three things kill it. The scale drifts as people score — a seven in row one is not a seven in row eleven. The arithmetic invites errors, and half the workbooks we have seen contain at least one broken formula. And it is tedious enough that clients skip it, which means the quarter starts with a wish list and a story about how it was scored.

ICE and its relatives

Impact, confidence, ease — three quick scores, multiplied. It is popular because it is fast and it forces a nod toward feasibility rather than pure desire.

The weakness is that multiplying three subjective one-to-ten scores produces a number that looks far more precise than its inputs. A 240 and a 216 are indistinguishable in reality, but the list will present them as a firm order, and people treat rank order as truth.

Pairwise comparison

Compare candidates two at a time, asking only which of these two matters more this quarter, and let the ranking emerge from the wins. Six candidates is fifteen comparisons and takes about four minutes; eight candidates is twenty-eight.

It works because it replaces one hard cognitive task — hold twelve items on an abstract scale — with many easy ones. Humans are good at binary comparisons and bad at absolute scales, and the method is built around that fact rather than against it.

It also produces the best artifact. When someone challenges the order later, you can show the specific comparison their item lost. That conversation takes thirty seconds instead of derailing a leadership meeting.

Where pairwise struggles

  • Large candidate lists. Comparison count grows quadratically, so cap the list at eight or so before you start.
  • Genuinely incomparable items. A compliance obligation and a growth initiative are not really rankable against each other — handle mandatory items separately, before the ranking.
  • Inconsistency. Humans produce circular preferences: A beats B, B beats C, C beats A. Good implementations surface this rather than hiding it, and the inconsistency itself is a useful conversation.

None of these are fatal, and all of them are manageable with a cap and a little facilitation.

What we recommend

Use pairwise for the ranking and keep weighted criteria as the tie-breaker underneath it. That combination gets the completion rate of the easy method and the defensibility of the rigorous one, which is the trade every coach actually wants.

And whichever you choose, do the ranking before the room, not in it. The purpose of the session is to challenge an order, not to build one under social pressure.

Where QuarterOS fits

QuarterOS runs pairwise comparison as step five of the guided flow, weighted by your own criteria, with the ranked list building live as the client decides — and the reasoning kept for the conversation three weeks later.

See how ranking works →
FAQ

Questions this raises.

How many candidates should go into the ranking?

Six to eight. Below six the ranking is barely worth formalising; above eight the comparison count climbs and clients start clicking rather than thinking. Cull the candidate list first — that cull is itself a useful exercise.

What about circular preferences?

They happen and they are informative. A circular result usually means two items are being judged on different criteria, which is worth ten minutes of conversation rather than being quietly resolved by software.

Can the leadership team each rank separately?

Yes, and the disagreements are often the most valuable output of the whole session. Where the CEO and the operations lead rank an item four places apart, there is a real misalignment worth naming.

Keep reading

Related reading

The feature this describes.

Where ranking fits in the agenda.

Where the candidates come from.

See what this looks like as software.

Fifteen minutes with the real product, and a straight answer on fit.

© 2026 QuarterOS. All rights reserved.PrivacyTermsSecurity