Trading Auditor
← back to every claim
INCONCLUSIVE

Bullish/Bearish engulfing

The engulfing shows one side taking control of the fight in a single stroke, and it is the most reliable reversal pattern in price action — with no indicator at all, only price.

Measured in forexEUR/USD and GBP/USD · daily bars built from 15m · 1 pip spread (~0.009% per round trip) · no survivorship bias

Net per trade
+0.04%
after fees
The fee is charged on both legs: every trade pays to open and pays to close.
Stalled at
costs
could not show the edge survives the cost
Worst drawdown
−0%
1,361 days underwater
How far the account fell below its own best previous moment.
Timeframe
1 day
The time grid this was measured on. The same technique on a coarser or finer grid is a different measurement, and can earn a different verdict.
  • invariance
  • costs
  • placebo
  • benchmark
  • out of sample
  • multiple testing
Equity curve
Risking 1.0% per trade, marked to market every day — not only when a trade closes. The line at 1.0 is the capital it started with.

At its worst the account was worth 0% less than its own best previous moment, and it spent 1,361 days below that peak.

The technique against chance
Same number of trades, same holding time, same assets — only the dates were drawn at random.
0%barthe technique+0.04%random dates−0.02%
The average trade and the typical one
The mean sits BELOW the median: most trades win a little and the tail loses a lot. It is the lottery inverted, and no less dangerous — what decides the outcome is the rare loss.
0%mean+0.05%median+0.06%

What was measured

We coded the rule exactly as it is described and let it trade on its own from 2020-01-01 to 2026-06-26. It found trades in 2 coins, 595 in total. Each one enters and exits at the price that existed on that day — the program never sees what comes next, which is the most common mistake people make testing strategies at home. Wherever the description was ambiguous we took the reading least favourable to the technique; whatever was left open is stated in the hypothesis, at the foot of this page.

How many of those trades actually count

Only 190 of the 595. Cryptocurrencies rise and fall almost in unison, so the same rule fires on dozens of coins on the same day — and that is one bet repeated, not dozens of different bets. Counting them all separately is the trick that makes a bad test look impressive.

What it paid, before any deductions

+0.05% per trade. This is the number the technique claims, and the only one on this page that has not yet been through a single control.

Where it stalled · the broker's fee

We subtract the fee the broker charges to open and to close each trade — every trade pays on both legs. What is left after the fee is too small to be told apart from zero.

What would have happened to the money

Risking 1.0% of capital per trade — the most widely taught rule — across the whole period: The capital would have ended at 1.00× what it started with. At its worst the account was worth 0% less than its own best previous moment, and it spent 1,361 days below that peak. In none of the 12 paths tested did it fall to less than half of what it started with.

What this result does NOT say

For an edge to be assertable here it would have to reach 0.65% per trade — that is the size that survives this archive's multiple-testing correction, and it rises as the archive grows. So INCONCLUSIVE means “we found nothing above that size”, and never “it cannot possibly work”. The instrument is far more sensitive than that: on synthetic data, with a clean effect, it separates from 0.09% upwards. The distance between the two numbers is the price of a real market and the price of publishing many claims. The difference matters, and it is the rule of this house: the card confronts the claim, never the person who made it.
What this card does not measure
Every verdict holds for the conditions it was measured under. These are this card's — and outside them the result does not apply.
One market, one universe
Measured on 2 spot currency pairs. It says nothing about futures, equities or indices, nor about how the same technique behaves in another market.
One window of time, not every window
The measured period runs from 2020-01-01 to 2026-06-26. A market moves through regimes, and a technique can work in one and fail in another — the card measures the regimes that fit inside this window, not the ones still to come.
One cost structure
The cost charged is a 1.0 pip spread, crossed once. Anyone paying more than that gets a worse result, and anyone paying less gets a better one — the verdict holds for this fee.
One exit rule
The trade was closed by: the auditor's fixed horizon (10 bars). The same entry measured with a different exit is a different strategy, and can earn a different verdict — it happens in this archive.
The number of trades is not the sample size
There are 595 trades, but only 190 independent market episodes: trades that overlap in time are not independent observations, and counting them as if they were inflates any result. It is the smaller number that governs the arithmetic. With 190 episodes, what the data supports is a range from -0.07% to +0.17% per trade — the published average is the centre of it, not the exact measurement.
Daily bars
Measured at the daily close. Nothing here measures what happens inside the day, and an intraday technique is not auditable with this data.

Numbers and reproducibility

The six controls

controlstatustthresholdepisodes
invariancepassed
costsinconclusive0.671.97190
placeboinconclusive1.001.97190
benchmarkpassed
out of sampleinconclusive
multiple testinginconclusive
  • invariance595 signals across 4055 barsconcentration: +30.55% of the profit sits in the top 5% of trades — with the dates shuffled, +30.17% (fails above +50.00%)
  • costsgross +0.049% · cost 0.008% · net +0.041% (t=0.67) · the range runs from -0.079% to +0.161%
  • placeboactual +0.041% · placebo -0.021% · excess +0.062% ± 0.062% (t=1.00 against a threshold of 1.97, 190 real groups, 29,750 sham dates, draw error ±0.008%)
  • benchmarktechnique +0.04% · buy and hold (same horizon) +0.01% · excess +0.03%
  • out of sampleasset half A: +0.069% (t=0.71, 160 episodes) · asset half B: +0.015% (t=0.20, 163 episodes) · liquid half (>= US$ 0/day): +0.041% (t=0.67, 190 episodes) · period 1/4 (2020-01-02 a 2021-07-25): +0.013% (t=0.10, 45 episodes) · period 2/4 (2021-07-30 a 2023-01-27): +0.052% (t=0.37, 46 episodes) · period 3/4 (2023-01-30 a 2024-08-26): -0.009% (t=-0.08, 47 episodes) · period 4/4 (2024-08-28 a 2026-06-09): +0.108% (t=0.98, 55 episodes)
  • multiple testing1 variation(s) tested · t=0.67 across 190 episodes (equivalent to t=0.67) · p≈0.5026 · false positives expected by chance ≈ 0.50

Equity — outside the six controls, and here is why

The t of the trade series is invariant to bet size: 0.5%, 1% and 3% agree to the sixth decimal. Nothing here moves the verdict — it moves what the account would have lived through.

risking 1.0% per trade
×1.00
worst drawdown from the peak
0%
days below the previous peak
1,361
signals refused for lack of capital
0%
paths where the account halved (out of 12)
0

Reproducibility

period
2020-01-01 to 2026-06-26
assets that traded
2
variations tested before this one
1
gross per trade
+0.05%
net per trade
+0.04%
exit rule
the auditor's fixed horizon (10 bars)
horizon bars
10
spread pips
1.0
episode days
12
seed
20260728

Twelve Data forex (15m aggregated to 1d) · collected on 2026-07-25 · 2 assets · 4,055 bars · 2020-01-01 to 2026-06-26

Hypothesis, filed before the result

A specific prediction: a candlestick pattern is the most fragile claim in the battery — an engulfing is a coincidence of four numbers. In crypto the placebo yielded almost the same as the technique. In forex I expect the same, and with one difference worth measuring: forex daily bars are AGGREGATED from 15m on a UTC boundary, not the market day that closes at 5 p.m. in New York — if a candlestick pattern depended on the boundary, it would be even more fragile here. I expect INCONCLUSIVE. Context that informs this prediction, declared so that it can be assessed: (a) the corpus has 101 claims in crypto, 4 of them passed — all from the moving average crossover family; (b) the available forex universe is only 2 pairs (EUR/USD, GBP/USD) over 6.5 years, against 540 pairs and 9 years in crypto, which gives ~88 trades against ~8,600; (c) the cost of turnover in forex is ~20× lower (1 pip ≈ 0.009% against 0.2% commission), so control 2 stops being the main killer; (d) I ran ONE smoke test on the 20/50 crossover and saw the result — see the disclosure in its own hypothesis.

filed on 2026-07-29, before the number existed

The original, as it was filed

Previsão específica: padrão de vela é a alegação mais frágil da bateria — um engolfo é coincidência de quatro números. Em cripto o placebo rendeu quase o mesmo que a técnica. Em forex espero o mesmo, e com uma diferença que vale medir: as velas diárias de forex são AGREGADAS de 15m com fronteira UTC, não o dia de mercado que fecha às 17h de Nova York — se padrão de vela dependesse da fronteira, ele seria ainda mais frágil aqui. Espero INCONCLUSIVO. Contexto que informa esta previsão, declarado para que ela seja avaliável: (a) o corpus tem 101 alegações em cripto, das quais 4 aprovadas — todas da família do cruzamento de médias; (b) o universo de forex disponível é de apenas 2 pares (EUR/USD, GBP/USD) em 6,5 anos, contra 540 pares e 9 anos em cripto, o que dá ~88 operações contra ~8.600; (c) o custo do giro em forex é ~20× menor (1 pip ≈ 0,009% contra 0,2% de comissão), então o controle 2 deixa de ser o principal matador; (d) ⚠️ rodei UM ensaio de fumaça no cruzamento 20/50 e vi o resultado — ver a divulgação na hipótese dele.

Pre-registration exists to keep prediction apart from rationalisation: written after the number, every hypothesis is right.

The same technique in the other market

The verdict held in the other market too, on independent data.

Earlier audits of the same technique

Each variation an author teaches enters as its own test, so that whatever might work in the strategy gets covered. The verdict held in all of them.

  1. this measurement →INCONCLUSIVE
  2. 2026-07-29INCONCLUSIVEopen ↗

record b28a2f449b49 · 2026-08-03 15:44

This code comes from this card's content: if anything here changed after publishing, the code would stop matching — that's how a change gets caught. We audit the technique, never the person — no record names an author, a channel or a brand.

The full record behind this verdict.