CoffeeScore methodology · version 0.2

A score has to mean something.

CoffeeScore is being designed as a category-aware, evidence-backed assessment—not a repackaged customer rating. Scores remain unpublished until a machine completes the required tests.

Current statusFramework published0 machines scoredTest handbook in pilot-draft state
01

Coffee performance

Taste, temperature, repeatability and category-appropriate drink quality.

02

Milk performance

Texture, temperature, speed and repeatability where the machine includes a milk system.

03

Ease of use

Setup, controls, workflow clarity and consistency for the intended buyer.

04

Cleaning

Daily effort, descaling, access to removable parts and avoidable mess.

05

Noise

Repeatable sound measurements during grinding, brewing and milk preparation.

06

Value

Performance and expected ownership cost against machines in the same category.

07

Build quality

Materials, fit, stability, parts support and service-oriented design.

08

Reliability

Warranty, failure evidence, repair pathways and long-term ownership data.

Category-aware weighting

Different machines should not be judged as if they do the same job.

Every row totals 100%. A zero means the dimension is not part of that category’s score.

CategoryCoffeeMilkEase of useCleaningNoiseValueBuild qualityReliability
Espresso25%15%12%10%8%12%10%8%
Bean-to-cup20%20%15%12%8%10%7%8%
Pod25%0%20%10%10%18%9%8%
Filter35%0%15%10%8%15%9%8%
Publication gates

No score until the evidence clears every gate.

Passing these gates permits publication; it does not guarantee a high score.

1
Retail-representative machine

Testing must use a normal retail unit with its model and firmware recorded.

2
Repeatable preparation

Core drinks are repeated under a published recipe rather than judged from one successful cup.

3
Measured evidence

Temperature, time, noise and energy fields must retain raw observations and test conditions.

4
Ownership cycle

Daily cleaning and the manufacturer’s descale or deep-clean process must be completed.

5
Commercial separation

Retailer commission, samples and manufacturer access cannot change a score or its weighting.

6
Visible limitations

Missing long-term reliability evidence must stay marked provisional instead of becoming a guessed number.

Scoring rules

Claimed, measured, provisional and tested are different states.

From framework to testing

The category test handbook is now public.

It defines shared controls, recipes, repetitions, measurements and the raw-data contract. Equipment and tolerances remain visibly draft until pilot testing validates them.