Judges
A crowd and a professional are different instruments
Free community rating threads give volume and noise; a commissioned judge gives one considered response, and people confuse the two constantly.
Guides on Judges: Boundaries belong to the judge, The case against the price list, What the skill in judging actually is
A community rating thread gives you volume from strangers who owe you nothing; a commissioned judge gives you one considered response from someone who agreed to look. They answer the same surface question and share almost nothing else, and confusing them is where most of the bad experiences in this space start.
What a rating thread actually is
Post a photo somewhere open, and anyone can reply. Some replies are careful, most are quick, and none of them owe you anything, because nobody agreed to be your judge - they showed up, said a thing, and left. There is no consent structure on the responding side beyond "this is a public thread," no accountability if a reply is cruel for its own sake, and no way to ask the specific person who wrote a comment to explain what they meant, because they are usually already gone.
Volume is the thread's actual product. Twenty strangers weighing in gives you a spread, which is a real kind of information - it tells you roughly where reactions cluster - but it tells you that at the cost of quality control on any individual reply, and the worst replies in a thread are frequently the loudest. Even the spread is fragile. Crowds can be genuinely wise - Lorenz and colleagues (2011) recall Galton's 1907 finding that a group's median estimate "can be more accurate than estimates of experts" - but in their own experiment with 144 people on simple estimation tasks, even mild social influence narrowed the diversity of opinions enough to undermine the effect. A thread where every reply is written after reading the ones above it is social influence by design.
What a commissioned judge is
A commissioned review starts with consent on both sides. You send a brief, a judge reads it and decides whether to take it, and the resulting clip is a considered response from a specific person who agreed to give you their attention, in the register you asked for. That is a transaction with a start, an end, and a professional on the other side of it who has a reputation attached to the name doing the work - which is a structurally different thing from an anonymous reply in a thread with no name attached to it at all.
The judge is also accountable in a way a thread commenter is not. A named professional with a persona and a portfolio has something at stake in the clip they deliver, and a bad or careless one shows up in how they are reviewed, which shapes whether they get commissioned again. Nobody in a rating thread is carrying that same weight, because nobody there is being paid or reviewed for the specific reply they left you.
Consent, accountability, register: the three places they diverge
Consent. A thread commenter agreed to browse a public forum, not to be your evaluator specifically. A judge agreed to look at your material because you asked and they accepted - the whole exchange happened on purpose, on both sides.
Accountability. A judge who misses a brief or delivers something careless can be disputed, and their standing changes because of it. A thread reply that lands badly just scrolls away, unattached to any consequence for whoever wrote it.
Register. You cannot ask a thread for a specific tone, and you get whatever twenty different strangers happen to feel like writing that day, unfiltered and unrequested. A judge performs the register you actually briefed, on purpose, which is the entire point of paying for one over posting for free.
Where the bad experiences come from
Almost every account of a rating experience going badly involves the free, public, anonymous version, not the commissioned one - a stray cruel reply in a thread with no context and no consequence, arriving alongside nineteen unremarkable ones, weighted by the reader as if it were the verdict rather than one anonymous opinion among many. That is not really a flaw in rating as an idea. It is what happens when you ask an unaccountable crowd a question that a considered, consenting individual would have answered more carefully - the instrument was wrong for the question, not broken.
A judge cannot fully replace what a thread gives you, either. If what you actually want is a wide spread of reactions rather than one, a thread does something a single commission cannot, and pretending otherwise oversells the commissioned version. The two are complementary instruments for different questions, not a cheap option and an expensive upgrade of the same thing.
What each is actually for
Want volume and a rough sense of where reactions cluster, and you can tolerate some noise in exchange for it costing nothing? A thread does that job. Want one person, chosen deliberately, responding to what you specifically asked in the tone you asked for, with something at stake in getting it right? That is what a commission buys, and it is worth knowing what actually arrives before you order one so the expectation matches the product.
If what you are chasing under either option is really a consistent number rather than an opinion of any kind, an algorithmic score gives you that without the crowd's noise or a judge's register. How that kind of tool reaches a number is a separate question from anything discussed here, and if the number you actually want is a measurement rather than anyone's opinion, a ruler settles it more honestly than either a thread or a judge ever will. For a straightforward comparison of the commissioned route against the algorithmic one as a purchase, Rate Cock has written that decision through directly, which is worth reading before you pick between any of these.