How to Build an LTL Carrier Scorecard

Build an LTL carrier scorecard with eligibility gates, five defined KPIs, fair comparison groups, sample-size rules, and a practical review process.

A blank five-row carrier scorecard on a warehouse worktable with two palletized loads behind it.

An LTL carrier scorecard turns shipment records into a repeatable carrier-selection tool. Start with eligibility, then measure each carrier on comparable lanes using a small set of clearly defined service, billing, freight-condition, and response metrics. Keep the raw counts beside every percentage so a perfect result from two shipments does not outrank a solid result from fifty.

The goal is not to crown one carrier as universally best. It is to identify which eligible carrier has performed well for the shipment profile you are about to quote.

Start with an eligibility gate, not a score

Before performance points are assigned, confirm that the carrier is eligible for the work. For interstate carriers, record the legal name and USDOT or docket number, then check current operating-authority and insurance information in the Federal Motor Carrier Safety Administration’s Licensing and Insurance system. Review the carrier’s FMCSA SAFER Company Snapshot separately. SAFER can show identification, operation type, cargo information, inspection and crash summaries, and a safety rating when one exists.

Treat those government records as screening inputs, not as a homemade safety grade or a substitute for the review your organization requires. Save the verification date because public records can change.

An eligibility worksheet can be simple:

Gate Evidence to retain Result
Identity Legal name, USDOT number, and docket number when applicable Pass, fail, or review
Operating authority Current authority record for the planned service Pass, fail, or review
Insurance filing Current FMCSA filing information relevant to the carrier Pass, fail, or review
Safety review Dated SAFER snapshot and any required internal review Pass, fail, or review
Service fit Written confirmation that the carrier handles the lane, commodity, equipment, and accessorial needs Pass, fail, or review

A failed gate should not be averaged away by a high service score. Escalate the record for review or exclude the carrier from that shipment’s comparison.

Define the comparison group before choosing metrics

Score like against like. At minimum, separate results by the dimensions that materially change the work:

  • origin and destination region or lane;
  • service type and delivery requirement;
  • commodity and packaging profile;
  • freight class or density profile when relevant;
  • accessorial requirements such as liftgate, limited-access, residential, appointment, or inside service; and
  • reporting period.

This matters because LTL rating and handling are sensitive to the freight presented. The National Motor Freight Traffic Association explains that NMFC freight class reflects density, handling, stowability, and liability. A carrier’s results for compact, easy-to-handle freight should not automatically predict performance for long, fragile, non-stackable, or otherwise difficult freight.

Use five metrics with explicit formulas

Choose definitions before looking at the results. The following structure is a practical starting point, not an industry-mandated benchmark.

Metric Formula What to record
Pickup reliability Qualifying on-time pickups divided by pickups with a documented pickup window Requested window, accepted window, actual pickup time, exclusion reason
Delivery reliability Qualifying on-time deliveries divided by deliveries with a documented delivery standard Standard used, actual delivery time, appointment or exception notes
First-pass invoice accuracy Invoices matching the accepted rate scope without an unexplained correction divided by audited invoices Quoted scope, invoice total, adjustment reason, responsible party
Freight-condition exception rate Shipments with a documented shortage or damage exception divided by delivered shipments Delivery notation, photos when available, claim reference, packaging notes
Exception response Exceptions receiving a substantive response within your written target divided by exceptions requiring carrier action Opened time, first substantive response, owner, resolution date

Define “qualifying” and “substantive” in the scorecard instructions. For example, a status acknowledgment is not necessarily a substantive response, and a missed pickup caused by freight not being ready should not automatically count as a carrier failure. Preserve the excluded shipment count and reason; silently removing unfavorable records makes the score unreliable.

Cost belongs beside these metrics, but it needs its own disciplined view. Compare the accepted quote, final invoice, and reason-coded differences. A higher final invoice is not automatically a billing error. Current carrier rules can allow charges or corrections for circumstances such as reweighs, reclassification, dimensional conditions, accessorial service, storage, or redelivery. Check the tariff or written pricing agreement that governed the shipment, and separate shipper-data errors from carrier billing errors.

Weight only what the team can measure consistently

Weights should reflect the business consequence of each metric and total 100%. One illustrative model is:

Metric Illustrative weight
Delivery reliability 30%
First-pass invoice accuracy 25%
Freight-condition exception rate 20%
Pickup reliability 15%
Exception response 10%

Convert each metric to a 0-to-100 score using a documented rule, multiply by its weight, and add the weighted values. Do not copy these weights blindly. A project-delivery operation may emphasize delivery reliability, while a fragile-freight program may place more weight on condition and claims handling.

Always display the denominator, reporting period, and missing-data treatment next to the weighted score. If a carrier has too few comparable shipments, show “insufficient data” rather than filling the gap with a neutral score. A weighted total without exposure counts creates false precision.

Work through a sample without inventing a benchmark

Suppose two eligible carriers served the same lane and service profile during the review period:

Carrier Comparable delivered shipments On-time deliveries Audited invoices Accurate on first review Condition exceptions
Carrier A 18 16 18 17 1
Carrier B 7 7 7 6 0

Carrier B has a perfect delivery result in this small sample, while Carrier A has more observations and a stronger invoice result. The right conclusion is not that one carrier is automatically superior. Show the calculated rates, flag Carrier B’s smaller sample, review the lane and shipment mix, and consider whether the next shipment resembles the records behind either result.

This example is illustrative only. Set minimum sample rules from your own shipment frequency and decision risk. When volume is sparse, use the scorecard as structured evidence alongside current quote terms, service fit, and documented exceptions rather than forcing a definitive rank.

Keep claims and damage data in context

Track at least three distinct facts: delivery exceptions, claims filed, and claims resolved. They are not interchangeable. A delivery notation may not become a claim, and an open claim should not be treated as a completed carrier outcome.

Also retain the governing documents. NMFTA’s Uniform Straight Bill of Lading terms address carrier liability, exceptions, and claim-filing requirements, while allowing a written agreement or carrier tariff to establish different or additional claim provisions. For each event, record the delivery date, notation, filing date, documents supplied, current status, resolution date, and outcome. Confirm the terms that actually applied instead of treating a general timeline as universal.

Review causes before changing the routing guide

A useful monthly or quarterly review asks four questions:

  1. Did each metric use the same definition across carriers?
  2. Were excluded records documented and applied consistently?
  3. Did a lane, commodity, packaging pattern, or accessorial requirement drive the result?
  4. Does the issue belong to the carrier, the shipper’s data, the consignee, or an unresolved handoff?

Discuss the underlying records with the carrier, assign corrective actions, and record an owner and review date. Keep current-period and trailing-period views side by side so one unusual shipment does not erase a longer pattern. When definitions or weights change, version the scorecard rather than rewriting history.

Use the scorecard during rate comparison

Apply the scorecard only after confirming the new shipment fits the comparison group. Prepare the origin and destination ZIP codes, packaged dimensions and weight, commodity description, freight class when known, ready date, delivery requirements, and needed accessorial services. Then compare the current rate and service terms with the carrier’s relevant lane-level record.

Shipocity is backed by a team with more than 40 years of combined logistics experience. Through established industry relationships, the platform helps businesses compare competitive freight rates for their specific shipment.

Compare rates for your next shipment

Sources