Scores
When two tools agree on the digit and nothing else
Two sevens from two tools are not the same seven; the agreement is on the display, not on the claim.
Guides on Scores: How much weight a rating tool result deserves, Every part of a rating result, and what each one is for, What can and cannot be compared, and how
No - two sevens from two rating tools are not the same seven. Each tool maps its internal judgement onto the scale by its own rule, so the matching digit is a coincidence of display, not agreement about the submission.
Same digit, different scale
Each tool's seven sits at a different point in that tool's own distribution. On a generous tool, seven might be roughly the median result. On a stricter one, seven might sit near the top of what it hands out. Where a tool centres its scale is rarely five, and rarely the same place from tool to tool - so a seven from a generous tool and a seven from a strict one describe two different positions in two different populations, dressed in the identical digit.
The two numbers being equal is not evidence the two tools rated the submission the same way. It is evidence that two different mappings happened to land on the same output value this time, which is a fact about the mappings more than about the submission. Even a change of wording on an otherwise identical scale moves the numbers: Steinberg and Rogers (2022) found that when every point of a response scale was labelled, the same personality and affect items drew different responses under agree-disagree labels than under not-at-all to very-much labels.
The reflex worth having
Before treating two matching numbers as agreement, ask where each seven actually sits: is it this tool's average, or its 90th percentile? Converting each number to its rank within its own tool, rather than comparing the raw digits directly, is the method that answers this properly when it matters enough to check.
For a quick read, even a rough sense of whether a tool runs generous or strict is usually enough to tell you the sevens are not equivalent, without needing the full conversion. A tool that shows its distribution, the way Rate Cock exposes its per-axis breakdown on public entries, makes that rough sense easy to get; a tool that shows only the number leaves you guessing at where its seven actually sits.
What this is not
This is not a claim that the two tools' underlying judgement disagrees - whether agreement between tools is meaningful at all is a separate question with its own answer, and depends on whether the tools share a backbone in the first place, which is knowable by checking how each model actually works. It is also not about a physical figure, where a matching number from two tapes genuinely would mean the same thing - that kind of agreement belongs to a stated method in a way a rating scale's output does not. And it says nothing about what two human reviewers scoring the same submission would agree on - that is a separate kind of number entirely, running on no shared scale with either tool.
A matching digit from two rating tools is worth a second look, not a nod of confirmation. The scale under each one is doing the work, and the scales are rarely the same scale.