Reviews
Out of scope for any clip
Comparisons to a population, anything about health, and predictions about other people's reactions are outside what a review can honestly contain.
Guides on Reviews: SPH as a genre with conventions, Five things a review is mistaken for, The words people use, defined, What happens in a two-minute clip
A review cannot tell you how you compare to everyone else, whether something is healthy or normal, or how other people will react. Those three questions sit outside what any judge can honestly answer, however skilled or well briefed, and knowing that before you write the request saves a disappointing answer or a polite decline.
"How do I compare to everyone else?"
A judge has seen the buyers who have commissioned them, and nobody else. That is not a population, it is a small and self-selected sample, and any answer built from it would be a guess dressed up as authority. A judge who takes this question seriously says so rather than inventing a percentile. If a population comparison is genuinely what you want, that is a measurement question with an actual method behind it, and it is the honest place to get that answer rather than a stranger's impression of where you might sit.
"Is this healthy or normal?"
A judge is not a clinician and has no business answering as if they were. Reassurance is a legitimate register and a judge can offer plenty of it, but reassurance about how something feels to look at is not the same as an answer to a medical question, and a professional judge stays out of that lane rather than improvising one. If there is a real worry behind the question, it belongs to a doctor, not a brief. In England, NHS talking therapies are free, accept self-referrals for anxiety and depression without a GP, and do not require a diagnosed condition. A related limit sits on the delivery side rather than the content: a judge will not promise how a clip will land emotionally, even when everything else in the brief is honoured.
"How will other people react to this?"
This one is asked more often than it looks, usually folded into a request like "be honest, would people like this?" A judge can tell you their own reaction, in the register you asked for. They cannot speak for a partner, a stranger, or people in general, because they are one person giving one considered response, not a survey. Two judges answering the same brief produce two performances, not two data points that average into a truth, and a single judge cannot stand in for a population any more than the score from an algorithm can, just from the other direction.
A fourth version worth naming
A close cousin of all three is "rank me against your other clients," which sounds like a population question but is really a confidentiality one. Answering it would mean discussing other buyers' material with you, and no professional judge does that regardless of how the request is phrased. This is one of the few briefs a judge will decline outright rather than answer in a softened form, because there is no honest version of the answer that does not involve breaching someone else's privacy.
Why the redirect is not a dodge
It is worth being explicit that a judge pointing you elsewhere for these questions is not evasion. A judge who invented a percentile, played doctor, or claimed to speak for a general audience would be giving you a worse answer dressed up as a more complete one, and a confident wrong answer is worse than an honest "that is not mine to tell you." The redirect is the accurate response, not a workaround for one.
Why these three, and not others
Each of these questions asks a review to be something it structurally is not: a dataset, a diagnosis, or a survey. A review is one person's considered, honest response inside a register, to a brief you wrote. That is a real and useful thing, and it is a specific-sized thing - what a human judgement can and cannot honestly cover has a shape, and these three questions sit outside it for the same underlying reason.
Compare this with what an automated score is built for, which is closer to a population question by design: consistency across many inputs is the actual product there, and it is a fair tool for a different job than a human judge is doing. Read correctly, that kind of score is a position in a distribution, which is exactly the comparison a human review cannot give you and an algorithmic one is built to.
Asking anyway
None of this means these questions are rude to have. They are common, and a good judge will redirect rather than refuse the conversation outright - toward a measurement, toward a doctor, or toward an honest "that is not something I can tell you." If you are unsure which instrument actually answers what you are after, Rate Cock's comparison of a human judge against an automated one is a better starting point than asking either to do the other's job.