Topic
Scores
What a number out of ten is built from, and what moves it.
-
A high total with a low axis, or the reverse
A breakdown that does not match its total is telling you about the weighting; a breakdown that contradicts itself is telling you about the noise.
-
A 7 that means "top 30%" and a 7 that means "7
Some tools map output onto their own past results and some onto a fixed rubric; both print out of ten and the numbers mean different things.
-
Why last year's 7 is this year's 7.5
A tool under commercial pressure has every reason to let its scale creep up and none to let it creep down.
-
What "objective" can and cannot mean here
A rating tool can be consistent, and that is the most it can be; objectivity would need a fact of the matter, and there is none.
-
A single number cannot carry the weight put on it
One rating is a draw from a distribution; the useful quantity is the centre of several, and the useful question is how many is several.
-
The line chart of your past scores, read carefully
A history graph mixes tool changes, photo changes and noise into one line, and the eye finds a trend in all of it.
-
Axis, defined
An axis is one component the tool scores separately before combining; the word promises independence it does not always deliver.
-
Ten is a convention, not a finding
The ten-point scale is inherited from everywhere else, and it imports assumptions the tools never checked.
-
The table people want, and why nobody can publish it
A chart mapping tool A's 7 to tool B's 8 would need both scales to be stable and both populations to be the same; neither holds.
-
Same tool, two submitters, two numbers
Even on one tool, two people's scores differ by their photos, their conditions and their devices before anything else; the comparison is mostly noise.
-
Using scores to order your own submissions
A tool orders a set of your submissions more reliably than it scores any one of them; the ends of the order are trustworthy, the middle is not.
-
Working out a tool's weights from its outputs
If a tool shows a breakdown and a total, a few results are enough to tell whether the total is a plain average or something else.
-
Distinguishing a difference from noise
A gap between two submissions matters when it is larger than the spread you see repeating either one; a rule of thumb, and where it fails.
-
The statistical reason retakes disappoint
An unusually high score is partly luck; the next one is expected to be closer to the centre, and that is not the tool changing its mind.
-
The bottom of the scale is mostly decorative
Rating tools rarely use their lowest values, and the reason is a design choice, not a compliment.
-
The conditions for a longitudinal comparison
Two scores months apart are comparable only if the tool, the conditions and the protocol were the same, and usually at least one was not.
-
What the written feedback in a result actually is
The prose is generated to match the number, not derived from an independent look; treat it as a caption, not a second opinion.
-
Six questions to answer first
Same tool? Same version? Same conditions? Same protocol? More than one sample each? Only then is the comparison worth making.
-
What the decimal place in a score is actually carrying
A second digit looks like precision; in most tools it is a rounding of an internal value that was never that stable to begin with.
-
A screenshot of a result, and what it leaves out
A shared score arrives without the tool version, the conditions or the retakes; it is a number stripped of everything that gave it meaning.
-
Bell, skew, or pile-up: what the histogram of a tool looks like
Two tools with the same average can distribute their scores very differently, and the shape decides what a given number is worth.
-
A 6.1 with glowing text, or an 8 with faint praise
When the paragraph and the score point in different directions, believe the score - the paragraph was written to it.
-
The confidence number, decoded
A confidence figure next to a score is usually about image quality or model certainty, not about whether the score is right.
-
Why extreme results circulate and ordinary ones do not
The scores people show are the best of many; the ones you see are selected, and selection is why they look impossible.
-
The population a score is implicitly compared to
Any score that means "better than most" depends on who "most" is, and tools rarely say.
-
When two tools agree on the digit and nothing else
Two sevens from two tools are not the same seven; the agreement is on the display, not on the claim.
-
The mean of two incompatible numbers
Averaging a 6.5 and an 8 from two tools produces a number on no scale at all.
-
The top of the scale, examined
A perfect score is a claim that nothing could move the number upward; almost no tool is built to make that claim.
-
The grade analogy, and why it misleads
Readers import school-grade meaning into rating scores; the two scales share digits and nothing else.
-
Two properties people conflate
A tool that gives the same photo the same score every time is consistent; whether the score is right is a separate question with no clean answer.
-
Composite, defined
A composite is a total built from parts by a rule; the rule is the part that matters and the part you are least likely to be shown.
-
The systematic tilt in rating tools, and where it comes from
A tool that scores low loses users; a tool that scores high keeps them. The centre of the scale moves accordingly.
-
Where a score gets rounded, and what that hides
A rounded score can move a whole point on a change too small to see; rounding is a threshold, and thresholds create false jumps.
-
A total and a breakdown are not the same information
A single number compresses several judgements and throws away which one moved; a breakdown keeps that, at the cost of being harder to brag about.
-
The comparative labels, decoded
A tool that says you are above average is comparing you to something; which something changes whether the label is worth anything.
-
Where a tool's standard comes from
A rubric encodes someone's decisions about what counts; those decisions are the standard, and no tool found them in nature.
-
The midpoint of the scale is not the middle of the results
Most tools centre their output somewhere between six and seven, and reading a six as below average is the most common misreading there is.
-
What the noun in "your score" refers to
A rating is a property of an image under conditions; the pronoun in "your 7.4" is doing a lot of quiet work.
-
What happens when many submissions crowd the top
If a tool scores generously, its top two points have to hold most of the population, and differences up there stop meaning much.
-
What the tool says it is scoring
Tools name the thing they score differently, and the name sets what a reader thinks the number is about; usually the number is the same underneath.
-
The most common score in the category, and why
Seven is where generosity, a crowded middle and a rounded display all land at once.
-
Getting from six numbers to one
Equal weights, fixed weights, and non-linear combinations all produce a total out of ten; each one makes a different submission look best.
-
What a score out of ten is made of
The scale is the most familiar part of any rating tool and the least examined. It is not a measurement, and it does not behave like one.
-
Agreement is weaker evidence than it looks
Two tools converging on the same score is reassuring and mostly uninformative if they were built the same way and given the same photo.
-
How much weight a rating tool result deserves
The number is real, the confidence is borrowed. This is the full argument for how much to trust a rating and where that trust runs out.
-
Every part of a rating result, and what each one is for
A result page has four or five distinct components, each produced differently and each deserving a different level of trust.
-
What can and cannot be compared, and how
Across tools, across time, across people, across photos, the complete map of which comparisons hold and what each needs to hold.
-
Reading a rating honestly
The number arrives with a confidence it has not earned. Six rules that put it back where it belongs.