The reading
A convention nobody chose
Four properties a reader needs in order to interpret a percentage are near-universally absent.
A numeric baseline is missing from 99.2% of claims, the source of the figure from 94.7%, the
period from 84.9%, and enumerated concurrent activity from 82.4%. They are absent at between 91%
and 99% in every vertical and at every tier of methodological capability. There is no subgroup in
this corpus where a reader is adequately served.
Four of the five conditions below cost a sentence. Stating what a percentage is calculated
from, over what period, alongside what other activity, and from what source requires no
experiment, no budget and no capability a firm does not already have. Only a comparison condition
requires doing more work. The costless ones are the ones missing, which is the evidence that this
is a convention gap rather than a capability gap.
What we require of a number
The conditions we apply before we will stand behind a figure, published because a paper about
disclosure should disclose its own. They are not proposed as a standard for anyone else to adopt.
- A numeric baseline. What the figure is a proportion of, in quantities a reader could
check.
- A comparison condition. What the result is measured against, or an explicit statement
that no comparison was constructed.
- A stated period, and whether it was fixed before the result was seen.
- Concurrent activity, enumerated, where several activities ran together.
- The source of the figure. Platform-reported, firm-calculated, or independently
produced.
What this study cannot say
Accuracy is not observed, and no claim here is that any published number is false. Published
pages are observed, not what a firm tells its client, and inconsistent disclosure does not imply
inconsistent rigor. The corpus is regionally and vertically weighted and is not representative of
the industry. Every claim in it was selected for publication by the firm that produced it, with
every incentive to demonstrate rigor, so every rate reported is a floor on non-disclosure rather
than an estimate of it. Automated classification carries measured error, reported per property in
the paper.
Two of the notes here work the same seam from the other side:
last-click is a reporting convention, not a
measurement method, and why your platforms disagree
with each other. Both are about numbers that carry a definition nobody states.
Competing interests. Dark Wave Marketing Science sells measurement services, so a study
finding that measurement is poorly disclosed is self-serving on its face. We raise it first
because it is the first thing a reader should ask, and we have tried to answer it by printing
the operative patterns in full, recording every failure found in the instrument, including one
that would have manufactured the expected result, and reporting the hypothesis that came back
null.
Data and code. The derived census records and the scripts that
reproduce every figure are available on request at
trevor@darkwavemarketing.science
and are being prepared for release. Raw crawled page text is not redistributed, but every claim
record carries its source URL, so any figure can be checked against the live page.