The 3PL Evaluation Checklist: 42 Questions Before You Sign
Choosing a 3PL often starts with a spreadsheet full of rates. That is also where many brands make their first mistake. One provider looks cheaper per parcel, another promises faster dispatch, and a third says its system can handle anything. Then onboarding begins, and the missing details appear: receiving fees, vague cut-off times, inventory adjustments without records, packaging decisions made on the warehouse floor, and claims that require evidence nobody knew to save.
A useful 3PL evaluation does not ask, ‘Who gave us the lowest quote?’ It asks, ‘Which provider can prove how the operation will work when orders, exceptions, and pressure arrive at the same time?’
Below is a copy-ready, download-ready scorecard built around eight areas: price, SLA, systems, inventory, packaging, claims, peak season, and exit terms. Give every question a score from 0 to 2: 0 means no answer or only a verbal promise, 1 means a partial answer without complete evidence, and 2 means a clear answer supported by a document, system record, or live demonstration. Any red flag should remain open regardless of the total score.
1. Price: compare the invoice, not the headline rate
1) Is every charge listed by unit, currency, and trigger? 2) Are receiving, inspection, storage, pick, pack, materials, relabeling, returns, and disposal priced separately? 3) How is dimensional weight calculated, and whose measurements control billing? 4) Which carrier surcharges may be added after dispatch? 5) How long are quoted rates valid, and how much notice is required before a change? 6) Can the provider produce a sample invoice based on your actual order profile?
Red flags include a cheap outbound rate with no accessorial schedule, ‘market price’ language without a notice period, and discounts that disappear above or below an unstated volume band. Ask for the current rate card, surcharge table, three anonymized invoices, and a worked invoice using your own SKU and destination mix. Price evidence should make the final bill predictable, not merely attractive.
2. SLA: define the clock before measuring speed
7) When does receiving time start: truck arrival, signed handover, or system scan? 8) What is the order cut-off time and time zone? 9) Are dispatch promises measured in working hours, calendar hours, or business days? 10) Which events pause the SLA clock? 11) What service credit or corrective action follows a miss? 12) How are normal and peak-season targets different?
A red flag is an SLA that says ‘same day’ but never defines a complete order, an operating day, or the system timestamp used as proof. Request a signed SLA schedule, a monthly performance report, timestamp definitions, exclusions, escalation contacts, and a sample corrective-action report.
For context, a BONDJET SLA sample can be expressed in measurable terms: normal inbound scanning within 2 hours and inspection photography within 4 hours; during September through December, inbound scanning within 4 hours and inspection within 8 hours. Oversized, abnormal, or unusually high-volume orders require separate confirmation. This is useful because the event, time window, peak exception, and operating boundary are visible. Your final SLA must still be confirmed for your product and workflow.
3. Systems: test the records you will depend on
13) Can the platform create SKUs, inbound forecasts, orders, and exception holds? 14) Does it keep an audit trail showing who changed what and when? 15) Can inventory and order data be exported without provider assistance? 16) Which integrations are native, and which depend on manual files? 17) What happens when an API, store connection, or carrier label service fails? 18) What user roles, access controls, and data retention rules apply?
Red flags include a polished dashboard with no export function, shared user accounts, edits without history, and integrations demonstrated only in slides. Require a live workflow demonstration, sample exports, field definitions, incident records, access-control screenshots, and a written recovery process.
4. Inventory: prove every movement
19) How are supplier parcels matched to the correct account and SKU? 20) Is expected quantity compared with received quantity? 21) How are damaged, unidentified, or short receipts quarantined? 22) How often are cycle counts performed? 23) What tolerance is allowed before investigation begins? 24) Can the provider show the chain from inbound scan to shelf, pick, pack, and dispatch?
A red flag is a warehouse that reports only one total stock number, corrects discrepancies silently, or cannot separate available, allocated, damaged, and held units. Ask for an inventory movement log, count policy, discrepancy report, quarantine procedure, and a sample SKU history. BONDJET supports inbound forecasts for larger or more complex receipts, SKU management, discrepancy updates, and inspection records, which gives a practical basis for a pilot demonstration.
5. Packaging: turn ‘careful handling’ into a specification
25) Who approves the packaging standard for each SKU class? 26) Are packaging materials, dimensions, and labor priced in advance? 27) Can the provider document packing with photos or records? 28) How are fragile, high-value, boxed collectible, and multi-part products handled differently? 29) When can staff substitute materials? 30) How are packaging changes tested against damage and dimensional weight?
Red flags include one packaging method for every product, unapproved substitutions, and no record of the condition before packing. Require a packaging SOP, bill of materials, sample packing record, photos, dimensional-weight calculation, and approval workflow. BONDJET can combine inspection photography, SKU checks, bubble wrap, EPE protection, reinforced cartons, wooden frames, or wooden cases according to product needs. The right choice must balance protection, weight, volume, and cost; reinforcement reduces risk but does not make damage impossible.
6. Claims: find out what happens after the parcel goes wrong
31) What events qualify as loss, damage, delay, or warehouse error? 32) What is the claim deadline? 33) Which documents and images must the seller provide? 34) Is compensation based on purchase value, declared value, selling price, or a fixed cap? 35) Which indirect losses and product categories are excluded? 36) Who owns carrier follow-up, and how often will status be reported?
Red flags include ‘we will help’ without a deadline, no published evidence list, and caps that are disclosed only after loss. Ask for the complete claims terms, a claim form, a resolved anonymized case, response targets, compensation caps, and insurance wording. BONDJET’s current process requires order or tracking details, a problem description, tracking history, and valid proof of value; damage claims also require product and outer-packaging photos. Exact protection and route terms should be confirmed before dispatch.
7. Peak season: measure capacity under pressure
37) Which months or events are treated as peak? 38) What volume forecast must the client provide, and by when? 39) What capacity is reserved versus shared? 40) Which SLA, carrier, labor, or surcharge changes apply?
A red flag is ‘we add staff’ without a staffing trigger, capacity figure, or priority rule. Request last peak’s throughput report, staffing plan, carrier allocation plan, forecast template, backlog escalation rules, and peak SLA.
8. Exit terms: protect continuity before signing
41) How will sellable, damaged, held, and unidentified stock be counted, packed, and transferred? 42) In what format and on what schedule will SKU, inventory, order, tracking, image, claim, and billing data be delivered?
Do not accept an agreement that says only ‘stock will be returned at the client’s cost.’ Require notice periods, final count rules, dispute handling, storage and handling fees, shipment priorities, data formats, retention periods, deletion confirmation, access to open claims, and named owners on both sides.
How to use the scorecard
Score all 42 questions, attach the requested evidence, and calculate each section separately. A provider scoring 70 out of 84 may still be unsuitable if inventory traceability or exit rights score zero. Treat missing evidence, not just low scores, as a decision item. Run references against comparable products and volumes, then move the best one or two providers into a controlled pilot.
For brands shipping high-value or complex-SKU products, BONDJET can be evaluated through the same evidence-first process. Ask the team to demonstrate inbound identification, inspection photography, SKU records, packaging approval, dispatch tracking, and exception handling with your own test products. The goal is not to collect confident answers. It is to leave the evaluation with records you can audit and terms you can operate.
