Agile estimation methods
Sprint estimation vs T-shirt sizing: which one should your team use?
Both methods exist to compare effort without pretending to measure hours. They differ in granularity, speed, and how far you can push the numbers into a forecast.
The short answer
Use T-shirt sizing when you are triaging a big, vague backlog and only need to know roughly how large things are. Use Fibonacci sprint estimation when the work is going into a sprint and you need numbers you can add up, compare across rounds, and turn into a velocity trend.
Side by side
| Dimension | Fibonacci sprint estimation | T-shirt sizing |
|---|---|---|
| Scale | 1, 2, 3, 5, 8, 13, 21 (or modified Fibonacci) | XS, S, M, L, XL |
| Granularity | High — the gaps force a real choice on bigger stories | Low — deliberately coarse |
| Speed | Fast per story, slower to reach consensus | Fastest for large volumes of items |
| Forecasting | Sums into sprint capacity and velocity | Needs a conversion step before it forecasts anything |
| Best team size | 3–10 people who own delivery of the stories | Any size, including cross-team roadmap sessions |
| Best stage | Refinement and sprint planning | Discovery, intake, quarterly roadmapping |
When T-shirt sizing wins
- You have 80 unrefined items and one hour. Numbers would be false precision.
- Stakeholders outside the delivery team are in the room and story points would invite arguments about hours.
- You are comparing initiatives, not stories — an epic is either an L or an XL and that is enough to sequence it.
When Fibonacci sprint estimation wins
- You need to commit to a sprint. "Three mediums and a large" does not tell you whether it fits; 3 + 3 + 3 + 8 against a 20-point average does.
- Complexity varies wildly inside one size bucket. The jump from 5 to 8 to 13 surfaces the disagreement a single "L" hides.
- You want a velocity trend. Points accumulate into a comparable number each sprint; letters do not.
- Estimates feed a re-vote. A spread of 2 to 13 is an explicit signal that the story is not understood yet.
The hybrid most teams land on
Size the roadmap in T-shirts once a quarter, then point the stories in Fibonacci during refinement. Anything that comes out XL gets split before it is pointed at all — if the team cannot agree between 13 and 21, the story is really two stories.
Whichever scale you pick, keep votes blind
The mechanism matters more than the scale. If the first person to speak sets the number, you are not estimating — you are ratifying. Everyone commits privately, then the facilitator reveals every card at once.
Stacked runs exactly that loop for Fibonacci and modified Fibonacci decks: the host shares one link, guests join with a name and no account, votes stay hidden server-side until the reveal, and a per-round timer keeps discussion honest.
Run a round in under a minute
Host a session, share the link, no signup for your teammates.
Stacked