Trading Auditor
← back to every claim
INCONCLUSIVE

Head and shoulders (and inverted)

Three peaks with the middle one highest mark the exhaustion of the trend; breaking the neckline confirms the reversal.

Measured in cryptoBinance spot · 540 pairs, delisted ones included · 0.2% per round trip

Net per trade
+1.56%
after fees
The fee is charged on both legs: every trade pays to open and pays to close.
Stalled at
placebo
the result depends on which draw came up
Worst drawdown
not measured
Equity was not measured for this claim.
Timeframe
1 day
The time grid this was measured on. The same technique on a coarser or finer grid is a different measurement, and can earn a different verdict.
  • invariance
  • costs
  • placebo
  • benchmark
  • out of sample
  • multiple testing
The technique against chance
Same number of trades, same holding time, same assets — only the dates were drawn at random.
0%barthe technique+1.56%random dates−0.15%
The average trade and the typical one
The mean sits above the median: the lottery signature — the typical trade loses and a rare handful pays for everything.
0%mean+1.76%median+0.03%

What was measured

We coded the rule exactly as it is described and let it trade on its own from 2017-08-17 to 2026-07-28. It found trades in 427 coins, 1,896 in total. Each one enters and exits at the price that existed on that day — the program never sees what comes next, which is the most common mistake people make testing strategies at home. Wherever the description was ambiguous we took the reading least favourable to the technique; whatever was left open is stated in the hypothesis, at the foot of this page.

How many of those trades actually count

Only 238 of the 1,896. Cryptocurrencies rise and fall almost in unison, so the same rule fires on dozens of coins on the same day — and that is one bet repeated, not dozens of different bets. Counting them all separately is the trick that makes a bad test look impressive.

What it paid, before any deductions

+1.76% per trade. This is the number the technique claims, and the only one on this page that has not yet been through a single control.

Where it stalled · against randomly drawn dates

We ran the whole thing again entering on randomly drawn dates and changing nothing else: same number of trades, same holding time, same assets. The gap between the technique and chance came out too small to assert anything. It is not that the signal is useless — it is that, with this much data, it cannot be told apart from drawing dates at random.

What this result does NOT say

For an edge to be assertable here it would have to reach 4.54% per trade — that is the size that survives this archive's multiple-testing correction, and it rises as the archive grows. So INCONCLUSIVE means “we found nothing above that size”, and never “it cannot possibly work”. The instrument is far more sensitive than that: on synthetic data, with a clean effect, it separates from 0.69% upwards. The distance between the two numbers is the price of a real market and the price of publishing many claims. The difference matters, and it is the rule of this house: the card confronts the claim, never the person who made it.
What this card does not measure
Every verdict holds for the conditions it was measured under. These are this card's — and outside them the result does not apply.
One market, one universe
Measured on 427 spot cryptocurrency pairs, delisted ones included. It says nothing about futures, equities or indices, nor about how the same technique behaves in another market.
One window of time, not every window
The measured period runs from 2017-08-17 to 2026-07-28. A market moves through regimes, and a technique can work in one and fail in another — the card measures the regimes that fit inside this window, not the ones still to come.
One cost structure
The cost charged is 0.10% per leg, in and out. Anyone paying more than that gets a worse result, and anyone paying less gets a better one — the verdict holds for this fee.
One exit rule
The trade was closed by: the auditor's fixed horizon (10 bars). The same entry measured with a different exit is a different strategy, and can earn a different verdict — it happens in this archive.
The number of trades is not the sample size
There are 1,896 trades, but only 238 independent market episodes: a single move fires the technique across dozens of assets at once, and counting those as separate observations inflates any result. It is the smaller number that governs the arithmetic. With 238 episodes, what the data supports is a range from +0.12% to +3.40% per trade — the published average is the centre of it, not the exact measurement.
Daily bars
Measured at the daily close. Nothing here measures what happens inside the day, and an intraday technique is not auditable with this data.

Numbers and reproducibility

The six controls

controlstatustthresholdepisodes
invariancepassed
costspassed1.881.97238
placeboinconclusive2.021.97238
benchmarkpassed
out of sampleinconclusive
multiple testingpassed
  • invariance1896 signals across 750934 bars
  • costsgross +1.762% · cost 0.200% · net +1.562% (t=1.88)
  • placeboactual +1.562% · placebo -0.154% · excess +1.716% ± 0.850% (t=2.02 against a threshold of 1.97, 238 real groups, 94,800 sham dates, draw error ±0.067%) — the status flips inside the placebo's own Monte Carlo error
  • benchmarktechnique +1.56% · buy and hold (same horizon) +0.20% · excess +1.37%
  • out of sampleasset half A: +1.162% (t=1.16, 211 groups) · asset half B: +1.979% (t=1.69, 192 groups) · liquid half (>= US$ 2,035,826/day): +1.850% (t=1.87, 219 groups) · illiquid half: +1.199% (t=0.89, 185 groups) · period 1/4 (2018-05-28 a 2022-08-08): +3.890% (t=2.48, 103 groups) · period 2/4 (2022-08-10 a 2024-02-02): -0.846% (t=-0.86, 53 groups) · period 3/4 (2024-02-04 a 2025-05-09): +3.927% (t=1.84, 41 groups) · period 4/4 (2025-05-10 a 2026-07-18): -0.646% (t=-0.48, 44 groups)
  • multiple testing1 variation(s) tested · t=1.88 across 238 episodes (equivalent to t=1.87) · p≈0.0620 · false positives expected by chance ≈ 0.06

Reproducibility

period
2017-08-17 to 2026-07-28
assets that traded
427
variations tested before this one
1
gross per trade
+1.76%
net per trade
+1.56%
exit rule
the auditor's fixed horizon (10 bars)
horizon bars
10
fee per leg
0.001
episode days
10
confirmation bars
5
tolerance
0.03
max window
120

Binance spot klines (delisted pairs included) · collected from 2026-07-27 23:31 to 2026-07-28 14:29 · 540 assets · 750,934 bars · 2017-08-17 to 2026-07-28

Seed not recorded: audits before 2026-07-29 used 20260728 by convention in the code, and the value is not in the record.

Hypothesis, filed before the result

Chart patterns are the best-known class in technical analysis and the one most damaged by the pivot error: almost every published backtest marks the signal on the pivot's own date, when on that date nobody knew it was one. Measured with a CAUSAL pivot — five bars of confirmation lag, which is what a real trader faces — I expect them to come out WORSE than the candlestick patterns, and for a legitimate reason: by the time the break is recognisable, the price has already moved. Specific, falsifiable prediction: none survives the Benjamini-Hochberg correction across the 10 patterns, and the control that kills the most will be the benchmark, not costs — unlike the candlesticks, because these trade far less often.

the filing date is not in this audit's record

The original, as it was filed

Padrões gráficos são a classe mais famosa da análise técnica e a que mais sofre com o erro de pivô: quase todo backtest publicado marca o sinal na data do pivô, quando naquela data ninguém sabia dele. Medidos com pivô CAUSAL — 5 velas de atraso de confirmação, que é o que o operador real enfrenta — espero que saiam PIORES que os de vela, e por um motivo legítimo: quando o rompimento é reconhecível, o preço já andou. Previsão específica e falsificável: nenhum sobrevive ao Benjamini-Hochberg de 10 padrões, e o controle que mais mata será o benchmark, não o custo — diferente dos de vela, porque estes operam bem menos vezes.

Pre-registration exists to keep prediction apart from rationalisation: written after the number, every hypothesis is right.

This claim has been audited once — there is no history to compare against.

record e8ca5d9d6281 · 2026-07-28 17:18

This code comes from this card's content: if anything here changed after publishing, the code would stop matching — that's how a change gets caught. We audit the technique, never the person — no record names an author, a channel or a brand.

The full record behind this verdict. No seed (predates 2026-07-29): the placebo may not replicate.