Choose My Stack
Join free
Editorial standards

How we compare software

16comparisons on this site publish scores out of 10. A score is worth nothing if you can't see where it came from, so this page explains exactly that — including what we don't do.

Rae Mercer, Editor, ChooseMyStackRae MercerEditor, ChooseMyStackUpdated September 2026
The scale

What a score out of 10 means

Every comparison scores both tools on criteria chosen for that category — ease of use and automation depth for project management, model choice and agent quality for AI coding tools. The criteria differ; the scale doesn't.

9–10Best in classGenuinely leads the category on this criterion. Hard to beat without changing what kind of tool it is.
7–8StrongDoes this well. Most teams will never hit its limits.
5–6AdequateCovers the basics. You'll notice the gaps if this criterion is your main reason for buying.
3–4WeakPresent but frustrating, or missing pieces most competitors include.
1–2Not a real optionEffectively absent, or only available on a tier most readers won't buy.
Where scores come from

What we do

Four things happen before a comparison is published.

01
We buy or trial every tool ourselves

Accounts are set up the way a reader would set them up, on the plan a reader would actually start on — usually the free tier or the entry paid tier, not a vendor-provided demo account.

02
We run the same tasks through each

Within a comparison, both tools get the same jobs: import the same data, build the same board, send the same campaign. Differences in the result are the comparison.

03
We read the full pricing page

Including the per-seat minimums, the annual-versus-monthly gap, the feature gates and the overage rates. Most of what makes software expensive is below the pricing table.

04
We check the vendor's own docs and changelog

Claims about what a tool can do are verified against its documentation, and scores are revisited when a changelog says something material shipped.

Limits and independence

What we don't do

This half matters more than the half above. Any site can claim it tests things; the useful question is what it refuses to do.

We don't run lab benchmarks

There is no test rig and no synthetic scoring. These are editorial judgements from using the products, and we'd rather say that than dress them up as measurements.

We don't score every edge case

A comparison covers what most readers decide on. If your requirement is unusual, the scores are a starting point, not an answer.

We don't accept payment for placement

No vendor has ever paid to appear, to rank higher, or to have a score changed. There is no mechanism for it.

We don't let affiliate links move a ranking

Some links earn a commission if you buy. The ranking is written before the links go on, and tools with no affiliate programme are recommended over ones that pay whenever they're the better answer.

Known limitations

Where these comparisons can be wrong

Three failure modes worth knowing about before you rely on a page here.

Prices go out of date

Vendors change pricing without notice. Every page carries the date it was last checked — treat any figure older than that as needing confirmation on the vendor's own site.

Scores are relative to the comparison

An 8 for “ease of use” means eight relative to the other tool on that page, not against all software everywhere. The same tool can score differently in two comparisons.

Fast-moving categories move fastest

AI tooling in particular can invalidate a verdict within a quarter. Those pages are re-checked more often, and the date at the top tells you when.

Corrections

How to challenge a score

If a score is wrong — a feature shipped, a price changed, or we simply misread how something works — say so and we'll check it. Corrections are made on the page itself and the updated date is moved, so you can tell the difference between a page that was re-checked and one that has been sitting untouched.

Vendors are welcome to write in on the same terms as readers. Being the vendor doesn't get a score changed; being right does.

Read a comparison

16 head-to-head and three-way breakdowns, each scored on this basis and dated with the last time it was checked.

All comparisons →