Cogitan
← Research

Programme · Active

Construct validity

Does the standard method measure what it claims?

Why this direction

Almost everything we have found that was worth publishing has this shape. Not "our model beats yours" — that framing hands the opponent the benchmark, and it is the one we have failed at repeatedly. What survives is narrower and more durable: a measurement in common use turns out not to track the thing it is used to decide.

These results need no competitor to be beaten, which is what makes them worth running. They are also uncomfortable to publish, because the same lens applied inward keeps finding our own numbers. We publish those too — a methodology critique from someone who exempts themselves is not a methodology critique.

Findings

8

Still open

What this direction has not answered, and what would settle each.

What must a device paper report for its claims to be independently checkable?

The median published device reaches 7–8 rules of our set. Saturated with what its own field already knows how to publish, it reaches 28 — so the binding constraint is reporting convention, not physics or tooling. Growing the database from 16 to 60 superconducting rules barely moved the number.

Settled if

A minimal reporting set lifts median reach from ~8 toward 28 across a large corpus — and closing it changes verdicts, not just counts.