Scores

A screenshot of a result, and what it leaves out

A shared score arrives without the tool version, the conditions or the retakes; it is a number stripped of everything that gave it meaning.

By Updated 4 min readScores

Guides on Scores: How much weight a rating tool result deserves, Every part of a rating result, and what each one is for, What can and cannot be compared, and how

A shared score tells you very little on its own: it usually arrives without the tool, its version, the photo conditions, the retakes or the breakdown, and its clean, confident frame hides that rather than declaring it. This is about reading that artefact honestly, not about the person who sent it.

What a human commissioning it might have said back is a separate exchange with its own etiquette, and it is not this one.

The tool is usually unstated

A 7.9 means nothing on its own; it means something relative to the scale it came from. A screenshot rarely names the tool, and even when a logo is visible, the version behind it is not. Two screenshots a year apart from the same service can be two different scales wearing the same interface, because rubrics get edited and mappings get retuned quietly. That drift is documented for AI services generally: Chen, Zaharia and Zou (2023) found GPT-4's accuracy at identifying prime numbers fell from 84% in its March 2023 version to 51% in June 2023, on the same questions. Without a name and a date, the number is a value with no declared unit.

The conditions are gone

Distance, angle, light, crop - the four variables worth fixing before a photo is submitted at all - are invisible in a result screenshot. A generous angle or a favourable crop can move a result by more than the difference between two people, and a viewer has no way to separate "the subject" from "the setup" in a single cropped number. The photo that produced the score is usually not even the photo attached to the share.

The retakes are gone

Almost nobody shares their first attempt. A shared 8.6 is frequently the top of several tries, not a single draw, and the gap between a first attempt and a selected best-of-five can be most of a point on a ten-point scale. Seeing one number tells you nothing about how many were taken to get it, and the scores people circulate skew toward the extreme end of what they actually saw for exactly this reason.

The breakdown, if there was one, is gone

A total-only screenshot has already discarded the one piece of information that would let you sanity-check it: which component moved. Even a tool with a real per-axis breakdown usually gets cropped down to the headline figure before it is shared, because the headline is the part built to travel. Some services keep that breakdown attached to a permanent public entry rather than a one-off image - Rate Cock is one, and a link to a public entry carries more than a cropped number ever will, because the axes and the submission conditions travel with it.

The rest of the set is gone too

Most people take several photos before choosing one to submit, and the tool often accepts more than one image if a set is offered. A shared screenshot shows the single result the sender kept, not the range across the attempts that did not make the cut, so a viewer sees the peak of a small search rather than a representative draw from it. Reading the highest score in circulation as if it were typical is the same error one level up: a single shared image is already the selected best of a set, before it is even the selected best of several tries.

What you can still infer

Not nothing. A shared score tells you the scale the person chose to submit to, roughly what register the tool speaks in, and - if you know the tool - approximately where its middle sits. It also tells you that the sender selected this result to show you, which is itself information: a result nobody would want to share is a result you never see, so the population of screenshots you encounter is already filtered upward.

What you cannot infer

You cannot infer the subject's standing on any absolute scale, because you do not have the conditions, the version, or the sample. You cannot infer that this was a first or typical attempt. And you cannot use it as your own benchmark - comparing your own result to someone else's shared number requires the same tool, the same conditions and more than one sample from each side, and a screenshot supplies none of those.

Read it as an artefact

The honest response to a shared score is not to distrust it outright, and not to take it at face value either. It is to treat it as a partial record - one number, one photo, one moment, selected - and to say so plainly if it comes up as a comparison. The camera work that produced it and the measurement, if any, that came with it are both upstream of the number and both invisible in the crop, which is the whole problem in miniature: a screenshot keeps the part that impresses and drops the part that would let you check it.

Read next

Full archive