Why Your Benchmark Differs
You measure your operation, compare it to a published figure, and the numbers do not match. This is the normal case and it usually means nothing about your performance. For another workplace-measurement reference, how Microsoft Teams tracks activity.
Seven reasons, and knowing them prevents both false alarm and false comfort. For complementary transport data and research, see Federal Maritime Commission.
The seven
One. A different definition. There is no standard one. A survey that did not define detention received whatever each respondent meant, and your careful gate-to-gate figure is not comparable to that blend.
Two. A different start event. Yours might be gate-in; theirs might be check-in. That gap is the yard, and it can be an hour.
Three. A different mix. Live load and drop-and-hook are different operations. An operation that is 70% drop will beat one that is 90% live, and neither is performing better.
Four. A different commodity. Refrigerated averages run substantially above dry — around 3 hours 16 minutes against about 1 hour 54 in the commonly cited 2021 figures.
Five. A different fleet size. Fleets of 25 trucks or fewer showed the highest dwell in that same data, at about 2 hours 23. Small operations get worse slots, and that is a market fact rather than a management failure.
Six. A different period. Rates, volumes and congestion move annually and seasonally. A 2021 figure describes 2021.
Seven. Self-report against measurement. Recall in a grievance context runs long. Your timestamps do not.
Any one of these can account for an hour. Several together can account for more than the figure itself.
What comparison is actually valid
Your facility against your other facilities, same period, same definition, same mix. This works and it is informative.
Your facility against itself, six months apart. The only comparison holding everything constant, and the one worth building a habit around.
And your lanes against comparable lanes on price — if freight into your dock costs more than similar distance and commodity into a neighbour, that difference is a benchmark the market computed, and it is more honest than any survey.
What to do when someone quotes a benchmark at you
Politely, and with the specifics.
What was the definition, what was the start event, and what mix does that cover? Ours is gate-to-gate on live loads, median two hours forty. If theirs is check-in to departure on a mixed book, we are probably not measuring the same thing.
That is not evasion. It is the accurate response, and it usually ends the comparison — because the person quoting the figure rarely knows the answers.
The trap in the other direction
Do not use incomparability as a shield.
If your median is four hours and the industry conversation is about two, the definitional differences do not close a gap that size. At some point the number is the number, and hiding behind methodology is the facility version of an undocumented invoice.
The test: compare your own periods. If your dwell is rising against your own baseline, no benchmark argument helps, and none is needed to know something changed.
What a useful internal benchmark looks like
Set it from your own best month, not from a survey. Median under ninety minutes, 95th under four hours, under 15% of visits over free time.
Per facility and per mode, because aggregating them recreates the comparability problem inside your own operation.
And review it against your own data quarterly, not against anyone's published figure.
That is a target you can defend, and it will be more demanding than any industry average because it is drawn from what you have actually achieved.
The short version
- Seven reasons your figure will not match a published one: definition, start event, mix, commodity, fleet size, period, and self-report
- Any one can account for an hour, and several together for more than the figure itself
- Valid comparisons: your facilities against each other, your facility against itself over time, and your lane prices against comparable lanes
- When a benchmark is quoted at you, ask for the definition, the start event and the mix — the person quoting rarely knows
- Do not use incomparability as a shield; if your median is four hours, methodology does not close that gap
- Set an internal benchmark from your own best month, per facility and per mode