How we score

Every score on this site comes out of one formula, and here it is. If you think we got a verdict wrong, you should be able to see exactly where the number came from and argue with it.

The formula

Three axes, each scored 1.0 to 10.0 to one decimal place. The overall is weighted, not an average:

overall = 0.35*Build Quality + 0.40*Reliability & Safety + 0.25*Value
AxisWeightWhat it measures
Build Quality35%Materials, hardware, tolerances, and how the thing survives weather and years outdoors.
Reliability & Safety40%Does it do its job every time, and what happens to your flock on the day it fails.
Value25%What you get for the money against the gear that actually competes with it.

Safety carries the most weight because this is gear where failure costs birds. A door that jams on a hen, a heat plate sitting against bedding, fencing a raccoon can reach through: none of that is redeemable by a good price. A cheap product can't score its way to the top on value alone, and that is deliberate.

The safety override

A product scoring below 6.0 on Reliability & Safety cannot ship with an overall above 7.0, whatever the weighted arithmetic returns.

The math is a tool, not the last word. When gear can hurt a bird on the way to failing, the verdict says don't buy it in plain words and points you at something better, and the number is capped so it can't contradict the verdict sitting next to it.

What the numbers mean

The same anchors apply to every axis and to the overall. We score against these, not against a private scale that shifts by category.

ScoreWhat it means
9.0 - 10.0Best in its category. No failure mode keepers report with any regularity. We would put it on our own coop without hesitation.
8.0 - 8.9Strong. Real caveats exist and we name them, but none of them change the recommendation.
7.0 - 7.9Solid with a genuine tradeoff. Right for some flocks, wrong for others, and the review says which.
6.0 - 6.9It works, with flaws serious enough that we point most readers at a better option.
5.0 - 5.9Only for a narrow case. Most people should buy something else.
Below 5.0Do not buy it. We say so plainly and send you somewhere better.

Expect a real spread. A catalog where everything lands between 8.0 and 9.5 means the person writing it was being polite, and polite scores are worth nothing to you. Some gear in this niche deserves a 5, and when it does, it gets one.

What goes into a score

We read the published specifications

Gauge and mesh size, motor ratings, dimensions, materials, temperature range, warranty terms. The claims a manufacturer makes are on the record, so we hold them to what they say.

We hold the product against what a flock actually needs

Floor space and roost length per bird, ventilation above roost height, a mesh a raccoon cannot reach through, an incubator that holds temperature closely enough to hatch. A coop that sleeps four hens is not a coop for eight, whatever the listing says.

We weigh long-term owner reports

Reddit's chicken-keeping communities and other keeper forums, where keepers post a year or three after buying. That is exactly when door motors, coop hardware and heated waterers start to fail, and it is the one thing a spec sheet can never tell you.

We verify every product is really buyable

A pick has to be a live listing you can order today, at a current-generation model number, from the brand we say it is. Rebadged imports and discontinued units get caught here.

What we don't do: hands-on testing

We have not held this gear. We don't run coops through a winter, put door motors on a bench rig, or weigh feed waste across a season. You will never read "in our testing" or "we measured" on this site, because it wouldn't be true, and a site that fakes that once has nothing left worth reading.

That limit is the reason our scores are auditable. Every point we give traces to something you can check yourself: a published spec, a dimension held against the space a flock needs, a linked post from a keeper three years into owning the thing, a live listing. Nothing rests on a measurement only we have seen.

It also means we say when something is unknown instead of filling the gap. "Nobody has run one of these through a hard freeze yet, so we're not scoring durability above a 7 until someone has" is a sentence we are happy to print. Naming precisely what nobody knows is more useful to you than a confident number with nothing under it.

Why store ratings never move a score

Amazon star ratings, farm-store ratings and reviews on a brand's own site are excluded from our evidence entirely. They are incentivized, they are gameable by anyone willing to pay for them, and the seller moderates them. They are the easiest signal on the internet to fake, which is precisely why we won't build a score on one.

What we use instead are places where keepers talk to each other with nothing to sell. Those threads are slower, messier and far more honest, and the criticism in them survives contact with a second winter. Where we quote a keeper, the quote is verbatim and links to the original post, so you can read the whole thread and decide whether we read it fairly.

How we're funded

Amazon affiliate links. If you buy through one, we may earn a small commission at no extra cost to you. As an Amazon Associate, we earn from qualifying purchases.

Some of our picks earn us nothing, and we say so on the page. When the best answer to a question isn't sold through any program we belong to, we name it anyway and link straight to the manufacturer. Telling you which links pay us is cheap, and it is the whole reason you should believe the ones that do.

No brand has ever paid for coverage, placement, or a better score, and none ever will. We decide what to recommend first and add the links afterward, never the other way around.

Think we got one wrong?

Tell us which score and why: corrections@chickengearguide.com. We read everything, and we correct in public. You can also read who writes this and why.