Options research · paper only · evidence gates
I test options ideas before they reach real money.
Every idea is written down, tested, checked, and kept on paper until the evidence is strong enough.
I built a paper-only system to challenge promising ideas. It records each question, tests it against history and controls, and stays away from real money until the evidence is strong.
An idea has to survive the whole loop.
Finding something interesting is easy. Knowing whether it is real is the job.
A historical test can produce a beautiful line and still answer a question nobody wrote down in advance. The danger is not that the result is ugly. It is that the result becomes persuasive before anyone asks what produced it by accident.
This project makes the challenge part of the system: finite questions, explicit controls, paper execution, and a record of what got rejected. Markets are the harshest place to learn that a persuasive chart can be nothing. That is the same lesson a campaign dashboard teaches, only faster and with a scoreboard.
Most ideas do not survive contact with a control group.
Open the detailed verdict distribution
The verdicts rest on a much larger evidence set.
15,412 historical-test results and 988,538 individual simulated-trade records fed into those 446 verdicts, over 6,479.2 hours of computer time. The verdicts are the conclusions; these are the records underneath them.
The research is only as trustworthy as the system running it.
A conclusion is only as good as the steps that produced it, so those steps are checked on a normal cadence—not just when a result looks surprising.
The useful result is the refusal to overread a pass.
The receipt behind the research result.
Open the research receipt
Counts and process metrics only. Strategy names, underlyings, thresholds, position sizing, and every positive edge value are deliberately excluded. The public page shows the checked summary; the raw working record remains private.
The claim
Verdicts recorded, the research gate, and the paper-only governance decision.
The checks
Controls, incidents, coverage, and the sanitized failure that stayed out of the ledger.
The scale
Backtest rows, matrix coverage, engine time, and operating cadence.
Let evidence decide.
The system is built to say "not enough evidence" without flinching, a habit worth having on any team.