route & fleet
Buying guides

Building a Vendor Evaluation Scorecard

How to score software vendors so the decision is comparable, defensible and not decided by whoever presented last.

Building a Vendor Evaluation Scorecard — illustration

Without a scorecard, software decisions are made on recency, presenter charisma and the preference of whoever is most senior in the room. A scorecard does not remove judgement; it makes the judgement explicit and reviewable.

Design principles

Weight before you look. Assign weights to criteria before seeing any vendor. Weights set afterwards describe your preference rather than your requirements.

Score evidence, not claims. A score above "partial" requires something you observed.

Score by section. Overall totals hide a vendor that is strong everywhere except the area you cannot compromise on.

Score independently, then discuss. Individual scores first, then reconcile. Group scoring converges on the first opinion voiced.

Record why. A one-line justification per score. Six weeks later, when the decision is questioned, the justifications are what defend it.

A feature promised for the next release is not a feature. Score it as partial with a note, and if it is critical, require a contractual commitment with a date and a remedy.

A workable structure

SectionWeightScored by
Planning and optimisation20%Planner, operations manager
Execution and dispatch10%Dispatcher
Driver mobile app20%Drivers, operations
Integration and technical12%IT
Reporting and analytics8%Operations, finance
Administration and security5%IT, compliance
Implementation approach and risk10%Project lead
Total cost of ownership10%Finance
Vendor viability and support5%Project lead

Adjust to your operation. A route accounting purchase would weight pricing and settlement heavily; a maintenance-focused purchase would weight the workshop module.

Note the driver app weight. In operations with many stops per day, the app is used thousands of times a week; the planning interface is used once. Weight accordingly, and let drivers score that section.

The scoring scale

ScoreMeaning
0Absent
1Partial — workaround required, or on the roadmap
2Adequate — meets the requirement
3Strong — meets it well, with evidence

Four points is enough. Ten-point scales produce false precision and endless argument about whether something is a 6 or a 7.

Mandatory gates

Separate from scoring, define pass/fail gates. A vendor failing any gate is eliminated regardless of total score:

  • Offline operation of the driver app for a full shift
  • Data export in a documented format at no cost
  • Support hours covering your operating hours
  • Required integrations demonstrably possible
  • Security certification your policy requires
  • Financial viability threshold

Gates prevent the outcome where a high-scoring vendor is selected despite failing something you cannot live without.

Running the scoring

  1. Circulate the scorecard with weights before the first demonstration.
  2. Each evaluator scores independently within 30 minutes of each session.
  3. Collect scores before discussion.
  4. Review divergences — where two evaluators differ by two or more points, discuss. Divergence usually means the requirement was ambiguous or one person saw something the other did not.
  5. Agree a consensus score with a recorded justification.
  6. Compute weighted totals by section and overall.
  7. Review the result. If it contradicts everyone's instinct, examine the weights — sometimes the weights were wrong, and sometimes the instinct was.

Presenting the decision

A one-page summary that survives scrutiny:

  • The weighted scores by section, all vendors
  • Gate results
  • Five-year total cost for each
  • The three most significant differentiators, with evidence
  • Key risks of the recommended option and how they will be managed
  • The recommendation and the reasoning

Attach the detailed scorecard. The summary is for the decision meeting; the detail is for the questions afterwards.

Common questions

How many people should score?

Four to eight, covering the affected functions, including at least one driver and at least one daily user of the planning system. Larger panels dilute accountability; smaller ones miss perspectives.

What if the cheapest vendor scores lowest?

That is the scorecard working. Present total cost of ownership alongside the score so the trade-off is explicit, and quantify what the score difference means operationally — usually in hours, failures or risk rather than in abstract points.

Should vendors see the scorecard?

Share the criteria and weights; withhold the scores. Transparent criteria produce better-targeted proposals and demonstrate a fair process, which matters if a decision is challenged.

How do we handle a tie?

Look at the section scores rather than the total, weight the sections that matter most in your operation, then use implementation risk and vendor viability as tie-breakers. A genuine tie usually means either option would work, in which case commercial terms and cultural fit should decide.

Can we change weights mid-process?

Only with a documented reason applied consistently to all vendors and re-scored. Changing weights after seeing results to produce a preferred outcome destroys the purpose of the exercise and will be visible to anyone who reviews it.

Nil Masferrer Jiménez · Editor · market and product research

Nil Masferrer Jiménez writes and edits Route & Fleet. His background is in business administration and finance, and the analytical spine of this site — cost per mile and per stop, total cost of ownership, payback and business-case models, software pricing structures and contract terms — is built on that. The operational and regulatory material is compiled from primary documentation: regulator publications, manufacturer and vendor technical specifications, and published industry research. Articles on compliance, telematics, maintenance and costs carry a Sources section linking those documents, so you can read the instrument itself instead of taking this summary on trust. He does not run a fleet, and the articles say so wherever that limit matters. Corrections are welcome and get published.

How this site is researched, and its limits

This article is editorially independent. Route & Fleet is funded by advertising displayed on the page; advertisers have no influence over its research or conclusions. See our advertising disclosure.

Keep reading

Related articles

Buying guides

The Demo Script That Exposes Weak Products

How to run vendor demonstrations on your terms — a scripted scenario approach that reveals what a standard demo is designed to hide.

19 August 2026 · 5 min read