The Top LTL Carrier Survey Needs a Lane-Level Reality Check

The latest LTL carrier survey has familiar winners. Old Dominion Freight Line earned the top national-carrier position for the 17th consecutive year, while Daylight Transport took the overall honor for a third straight year, according to FreightWaves' report on the Mastio survey. That consistency says something important about service quality and customer loyalty.
It does not, however, answer the question a transportation manager must answer on Monday morning: Which carrier should receive the next shipment from this origin to this destination, with this product profile and delivery requirement?
The difference matters. A national ranking aggregates thousands of experiences across networks that are inherently uneven. LTL freight moves through specific terminals, relay points, linehaul schedules, delivery routes, and appointment processes. A carrier can be excellent nationally and merely average on one terminal pair. A regional carrier can be exceptional in its core footprint yet become less competitive when freight crosses into an interline or distant service area.
Surveys are a valuable starting point. Allocation decisions need a lane-level reality check.
What a national survey can—and cannot—tell you
Reputation measures are useful for identifying carriers worth evaluating. They can reflect broad strengths in reliability, sales relationships, technology, claims handling, and perceived value. Long winning streaks are not accidental; they generally indicate that a carrier has built repeatable processes across much of its network.
But an aggregate score blends together customers with different freight classes, shipment densities, accessorial needs, geography, and definitions of “on time.” It may also disguise local variation. The terminal that collects a shipment, the breakbulk facilities it passes through, and the final delivery terminal can influence the outcome more than the logo on the trailer.
Even category scores need context. Logistics Management's Quest for Quality reporting, for example, separates dimensions such as on-time performance, information technology, customer service, equipment and operations, and value. In its 2023 LTL results, Old Dominion led four of those categories, while R+L Carriers led value. The lesson is not that one measure is superior. It is that “best” changes with the outcome a shipper values.
Build the scorecard at the terminal-pair level
Begin with origin and destination terminal pairs, not broad regions. Then segment the data by shipment profile: standard palletized freight, long or fragile goods, high-density freight, appointment deliveries, residential or limited-access stops, and other meaningful operating categories.
For each segment, track a compact set of metrics:
- Pickup reliability: pickups completed within the agreed window, with shipper-caused misses identified separately.
- Transit reliability: deliveries made by the carrier's quoted date and time commitment.
- Appointment performance: appointments requested promptly, kept as scheduled, and documented when the receiver changes them.
- Scan integrity: expected pickup, terminal, out-for-delivery, and proof-of-delivery events present and timely.
- Exception-free delivery: shipments delivered without shortage, damage, refusal, or service exception.
- Claims performance: claim frequency, dollars claimed per freight dollar, resolution time, and recovery percentage.
- Invoice accuracy: invoices matching the contract, shipment attributes, and approved accessorials.
- Recovery: elapsed time from an exception signal to a credible new plan and final resolution.
This approach aligns with established practice. Inbound Logistics identifies five core carrier-scorecard measures: pickup performance, on-time delivery, exception-free delivery, claims-free service, and invoicing accuracy. Its more recent guidance also recommends analyzing each carrier by key lane and shipment type before adjusting the carrier mix.
Join the records before judging performance
No single dataset tells the full story. Tender records show acceptance and promised service. Carrier scans show physical progress. Delivery records establish arrival and proof of delivery. Audit data exposes billing accuracy. Claims records reveal damage severity and resolution behavior.
CXTMS can bring those events into one shipment timeline and apply consistent definitions. That consistency prevents three common distortions.
First, a delivery scan alone cannot establish whether a load was on time unless it is compared with the accepted commitment and appointment history. Second, a low claim count can be misleading when claims remain open or recovered dollars are poor. Third, a cheap linehaul rate is not cheap after avoidable accessorials, reclassification charges, re-delivery costs, and internal exception labor are added.
Data hygiene is part of the scorecard. Require a minimum shipment count before ranking a lane. Display sample size beside every percentage. Exclude cancellations and shipper-caused delays through documented reason codes rather than informal judgment. Use rolling periods long enough to smooth random variation but short enough to expose a terminal that is deteriorating now.
Turn scores into controlled allocation changes
A scorecard should change behavior, not decorate a quarterly review. Establish a baseline allocation for every meaningful lane and define thresholds in advance. For example, a carrier that remains above the lane's on-time target and below its claims ceiling may earn additional share. A carrier that falls below threshold could enter a review period before losing volume. Critical shipments may require stricter gates for scan completeness and recovery speed.
Weight the score to the service promise. A retail replenishment lane may prioritize appointment compliance and exception recovery. Industrial freight with difficult dimensions may put more weight on damage-free handling. Routine replenishment may emphasize consistency and total landed transportation cost.
Avoid reacting to one bad shipment or moving an entire network after one survey. Shift allocation in measured increments, monitor the next operating window, and keep a capable secondary carrier active. The goal is not to crown a permanent winner. It is to create a feedback loop in which carriers earn freight through repeatable outcomes on the lanes where they operate.
National surveys remain useful evidence. They reveal durable reputations and give sourcing teams credible candidates. The stronger decision combines that external signal with the shipper's own tender, scan, delivery, invoice, and claims history. That is how an admired carrier becomes the right carrier for a specific shipment—or does not.
Ready to replace national averages with lane-level LTL decisions? Request a CXTMS demo to build carrier scorecards from your own operating data.


