Tools
The scorers that came before image tools
Before tools looked at photos, they scored questionnaires; the habit of a number out of ten with a paragraph of prose comes from there.
Guides on Tools: Every presentation choice on a result page, and what it does, A taxonomy of rating tools by what they output, A method for judging any rating tool before trusting it
Before rating tools looked at photos, they scored questionnaires: you answered questions about yourself, and a fixed set of rules turned the answers into a number. What that era left behind is the format - a score out of ten plus a paragraph of commentary - which image tools inherited without re-deriving it.
What the format actually was
No photo was involved anywhere in the pipeline. A quiz-era result was a function of self-report, run through a fixed set of rules that mapped answers onto a score. It was, structurally, closer to a personality test than to anything a camera would recognise. What made it feel like a rating tool rather than a survey was the packaging: a number out of ten, a short paragraph of generated commentary explaining what the number meant, and a page built to be screenshotted and shared.
What outlived the questionnaires
When image models became cheap enough to run this kind of tool on an uploaded photo instead of a set of answers, the underlying mechanism changed completely. The packaging did not. The number-out-of-ten-plus-paragraph format that quiz sites had already normalised became the default shape for image-based tools too, without anyone re-deriving it from what a vision model actually produces. The image side had been in the research literature for years by then: Eisenthal, Dror and Ruppin (2006) trained a predictor on human beauty ratings of face photos and reported a correlation of 0.65 with the average human rating. That is one reason the prose under a modern score reads like commentary rather than analysis - it is filling a slot the format created before there was anything visual to analyse.
The habit of a single total, rather than a breakdown, also traces back further than the current models. Quiz-era tools mostly returned one number because a questionnaire naturally reduces to one score; image tools inherited the single-total default before multi-axis reporting arrived as a separate, later decision.
Why this matters now
Knowing a convention is inherited rather than derived changes how much weight to put on it. The number-plus-paragraph shape was not designed around what an image model can actually tell you about a submission; it was already the house style of the category before that model existed. The fuller lineage from public voting through quiz sites to current tools traces where the rest of today's format came from, including the parts that predate this era too.
None of this touches what a current vision model is actually computing when it scores a photo, which is a genuinely different question from where the display format came from. It also has no analogue in a measurement taken with a tape, which was never packaged this way, or in a written response from a human reviewer, which owes nothing to the quiz-site format at all. Rate Cock still delivers a number and prose in that inherited shape, alongside the six-axis breakdown that the format itself never required.