The simplest guard against a finding that is really one good week: cut the record in two by date and require the effect to hold in both halves.
Sweep enough data and you will find patterns. Most are one good week wearing the costume of a rule. The half-split test is the cheapest way Apex knows to tell them apart.
Cut the record in two by date — not by row count, so that one heavy day cannot sit in both halves and test itself. Measure the effect in each half against the unconditional baseline. A finding passes only if the sign holds in both.
It does not correct for multiple testing. Apex logs every hypothesis ever tested and judges significance against the whole family (a Bonferroni correction): when 276 tests produced 27 with p < 0.05, that was roughly what chance alone produces, and one finding survived.
See also: methodology · what is decision intelligence.
All numbers on this page were published in Apex’s own record at the time; they describe the past under stated conditions and promise nothing about the future. All articles · The Apex day