Do you think a numeric score should mean the same thing across completely different product categories, or only within the category being reviewed?
I lean toward consistent definitions for the criteria but category-specific weighting. Otherwise a strength that matters a lot in one type of product can distort the score somewhere else.
Side-by-side notes or it did not happen.

