Quant Memo
Core

A Library of Negative Results

A research governance practice of formally recording every idea that was tested and rejected, so a team doesn't waste months re-testing the same failed hypothesis and can tell how much fishing produced a surviving signal.

Quant research teams generate far more failed ideas than successful ones, but most shops only keep records of what worked. Without a record of what was tried and rejected, researchers re-test the same signal variants every year or two as staff turn over, and — more seriously — nobody can tell how many hypotheses were quietly discarded before a "successful" one surfaced, which is exactly the information needed to judge whether that surviving signal is real or just the winner of a multiple-testing lottery.

A negative results library is simply a maintained log: every hypothesis tested, its specification, the data and period used, and the outcome, kept regardless of whether it passed or failed. Two things fall out of keeping it. First, it prevents duplicated effort — a quick search before starting new research shows whether a variant of the idea was already tried and why it failed. Second, and more important statistically, it lets a reviewer compute an honest multiple-testing correction: if the library shows 200 related hypotheses were tested before this one survived, the true significance of the survivor needs to be adjusted downward accordingly, something impossible to do if the 199 failures were never recorded.

Logging every tested-and-rejected hypothesis, not just the winners, prevents duplicated research effort and — more importantly — supplies the count of prior attempts needed to correct a surviving signal's significance for multiple testing.

Related concepts

Practice in interviews

Further reading

  • Lopez de Prado, Advances in Financial Machine Learning, ch. 8
ShareTwitterLinkedIn