EA Verdict audit: The Gold Reaper MT5

GOLD.ls · M1 (2020.01.01 - 2026.08.11) · initial deposit 10000 · leverage 1:100 · engine 0.1.0 · schema 0.1

Verdict summary

Data qualityStructureCostsConcentrationRegime dependenceProp-firm fit
INFOCAUTIONINFOCAUTIONOKCAUTION
Heavy stacking: up to 16 simultaneous positions. 80% of profit made in 31 days. [iqcapital_classic] Survives only at 1/4 sizing.

What we measured

Trades reconstructed4258
Pairing confidencemixed:log+fifo (71% log-exact)
Balance reconstructionexact
History quality39% real ticks
Real ticks from2024.01.02 00:00:00
Execution delay10 ms
Tester log providedyes
Tester log files20260811.log, 20260812.log
Randomizer prints0
Report SHA-25667873c23878e30ba…

Reconstructed floating equity

Floating drawdown (sampled)22.9 %
Balance drawdown (deal-wise)22.1 %
Report-head equity DD (tick-based)25 051.59 (9.24%)
Equity samples (deal timestamps)8516

sampled at deal timestamps; intrabar floating equity between deals is unobservable from a report, so the tester-head equity DD (tick-based) is the upper reference

Findings

INFO Pairing log-exact for 71% of trades
Exits confirmed by the tester log are paired exactly; the remainder (manual/basket closes, which print no trigger line) falls back to FIFO. Aggregate figures are unaffected; per-strategy attribution of the FIFO share may be imprecise.
INFO Only 39% real ticks
Most of the window uses generated ticks; fill realism is limited before the real-tick start date.
CAUTION Heavy stacking: up to 16 simultaneous positions
3392 stacked entries: 723 split tickets (one signal, several tickets), 2164 adds with the move, 494 against it. No dominant averaging geometry, but margin use and prop-rule exposure scale with the stack.
INFO Long and short held simultaneously
25 moments with open positions on both sides.
INFO Financing costs dominate
Swap (-43083) exceeds commission (-10654): the system holds positions overnight and over weekends, and pays for it.
CAUTION 80% of profit made in 31 days
That is 3.5% of 882 trading days. Miss a handful of days and the edge is gone; consistency rules at prop firms punish exactly this profile.
CAUTION [iqcapital_classic] Survives only at 1/4 sizing
Death-free first at 1/4 of the tested size, leaving ~12.99%/a of withdrawal. The marketed returns are not reachable inside this rule set.
CAUTION [iqcapital_classic] Position-loss limit breached even at 1/3 sizing
410 trades exceed the per-position loss limit at one-third size (767 at full size).
CAUTION [iqcapital_classic] Consistency rule breached in 2 year(s)
CAUTION [generic_6pct_trailing] Survives only at 1/4 sizing
Death-free first at 1/4 of the tested size, leaving ~12.99%/a of withdrawal. The marketed returns are not reachable inside this rule set.

Sub-strategies

Entry-comment prefix: "The Gold Reaper_XAUUSD_". The split is based on entry comments; MT5 tester artifacts do not carry magic numbers, so magic-only multi-strategy EAs appear as one group here (behavioral clustering is on the roadmap).

KeyTradesPnLWin %Long/ShortMedian hold (h)
810613419229.5573/4880.51
68262612166.2443/3830.58
58153242171.8452/3630.5
77262930570.2347/3791.01
43654762450.1197/1683.52
31853118152.4127/581.12
11563325657.1123/333.0
21243748063.786/382.58

Underlying vs. strategy

SymbolTradesEntry-price moveTraded spanLong/Short tradesLong PnLShort PnL
GOLD.ls4258+184.1 %2020-01-02 … 2026-08-102348/1910166340105239

approximated from entry prices; a tester report carries no independent price series. Price-percent and account-percent are not directly comparable without sizing; the move contextualizes the long/short split, it is not a benchmark return.

Year by year

YearTradesPnLWin %Long PnLShort PnL
2020598265553.02848-193
2021711369850.114832215
2022664125856.3-8182076
20236721822857.3132964932
20246473924256.12858010662
20256488044860.35882721621
202631812605068.26212463926

Long vs. short by year, against the underlying's own move. Where the dark bars rise and fall with the red ones, the market's tailwind is doing the work:

63,9260+63%0%2020 long_pnl: 2,8482020 short_pnl: -1932020: +24.6%20202021 long_pnl: 1,4832021 short_pnl: 2,2152021: -4.7%20212022 long_pnl: -8182022 short_pnl: 2,0762022: +1.3%20222023 long_pnl: 13,2962023 short_pnl: 4,9322023: +13.0%20232024 long_pnl: 28,5802024 short_pnl: 10,6622024: +28.1%20242025 long_pnl: 58,8272025 short_pnl: 21,6212025: +62.8%20252026 long_pnl: 62,1242026 short_pnl: 63,9262026: -2.8%2026top: long PnL = dark, short PnL = light (USD) · bottom: underlying move per year (entry-price approx.)

Costs and honest metrics

Net profit271578.96
Commission-10654.00
Swap-43083.44
Gross before costs325316.40
Cost share of gross16.5 %
EOD Sharpe (annualized)1.86
Report-head 'Sharpe'4.14
CAGR62.7 %
Max drawdown (EOD)22.0 %

The report-head "Sharpe" is trade-based and not comparable to an annualized daily Sharpe; the EOD figure above is the honest one.

Cost fragility and outlier dependence

Break-even cost shock: +63.78 USD per trade. Add that much cost to every trade (worse spread, slippage, commission) and the whole result is gone. Average position size: 0.25 lots.

Extra cost per tradeNet profitProfit factorWin %
+0.50 USD2694501.93156.3
+1.00 USD2673211.9255.7
+2.00 USD2630631.953.6
+5.00 USD2502891.83848.8

Leave-best-N-out: remove the N most profitable trades:

Best trades removedTheir PnLShare of gross winsNet without them
1134802.4 %258099
55703410.2 %214545
108263714.8 %188942
2011238420.1 %159195

flat USD shock per trade, not lot-scaled -- judge it against the average lot size; leave-best-out removes the N most profitable trades from the observed set.

Exit profile

Exit typeCount
SL2116
TP1281
signal_or_time861

Prop-firm fit: IQ Capital Classic (funded)

Rules used for this simulation (as of 2026-08-13):

dd_modeeod_trailing
dd_pct6.0
max_position_loss_pct0.5
consistency_pct30.0
overnight_allowedTrue
weekend_allowedTrue

Source: https://support.iqcapital.io (interne Extraktion docs/architecture/prop_iqcapital_regelwerk_2026-08.md)

SizingAccount deathsWithdrawal %/a
1/13693.94
1/1.51853.76
1/21035.6
1/3320.03
1/4012.99
1/608.7
1/806.54

Max-position-loss breaches (per sizing): 1/1: 767, 1/1.5: 672, 1/2: 567, 1/3: 410, 1/4: 317, 1/6: 235, 1/8: 182

YearBest dayYear profitBest-day share %Breach
2020955265536.0YES
20211011369827.3no
20221193125894.9YES
202331341822817.2no
202461153924215.6no
202587238044810.8no
20261743612605013.8no

Worst days: 2026-04-23: -10355, 2024-10-10: -6917, 2025-10-13: -6125 · overnight trades: 587 · weekend-spanning: 158

Prop-firm fit: FTMO Challenge

Rules used for this simulation (as of 2026-08-13):

dd_modestatic
dd_pct10.0
daily_loss_pct5.0
overnight_allowedTrue
weekend_allowedTrue

Source: https://ftmo.com/en/how-it-works/ + academy/maximum-daily-loss + faq weekend

SizingAccount deathsWithdrawal %/a
1/11678.17
1/1.5441.03
1/2127.24
1/3017.23
1/4012.99
1/608.7
1/806.54

Daily-loss breaches (per sizing, EOD deltas): 1/1: 114, 1/1.5: 75, 1/2: 58, 1/3: 40, 1/4: 33, 1/6: 18, 1/8: 8

Daily-loss breaches, intraday equity (per sizing): 1/1: 164, 1/1.5: 108, 1/2: 76, 1/3: 54, 1/4: 37, 1/6: 24, 1/8: 12. intraday row: sampled equity vs. day anchor (server-day boundary); EOD row for comparison

Worst days: 2026-04-23: -10355, 2024-10-10: -6917, 2025-10-13: -6125 · overnight trades: 587 · weekend-spanning: 158

Prop-firm fit: Generic 6% EOD trailing

Rules used for this simulation (as of 2026-08-12):

dd_modeeod_trailing
dd_pct6.0
overnight_allowedTrue
weekend_allowedTrue

Source: generisch, keine Firmenquelle

SizingAccount deathsWithdrawal %/a
1/13693.94
1/1.51853.76
1/21035.6
1/3320.03
1/4012.99
1/608.7
1/806.54

Worst days: 2026-04-23: -10355, 2024-10-10: -6917, 2025-10-13: -6125 · overnight trades: 587 · weekend-spanning: 158

Monte-Carlo resampling: stress on the reconstructed trades

1000 paths per method, seed 42, block length 5 trading days; additive resampling at the tested sizing; drawdowns as % of start balance. Paths are not stopped at account death -- drawdowns beyond 100% mean repeated wipeouts at this sizing.

Max-drawdown distribution (percent of start balance):

MethodMedianP90P95P99
Permutation (order only)83.37 %111.58 %121.54 %147.8 %
Bootstrap (IID)81.52 %115.64 %129.88 %153.38 %
Block bootstrap (daily blocks)172.29 %252.92 %289.51 %368.94 %

Observed EOD max drawdown of this backtest: 22.0 %. Compare it against the percentiles above.

Equity fan: cumulative PnL of resampled paths vs. the observed backtest (% of start balance):

3741%0%-60%1722 trading daysobserved = red, median = dashed, bands = P25-P75 / P5-P95

Bootstrap intervals (5th … 95th percentile):

MetricP5MedianP95
Profit factor1.6991.9452.234
Expectancy per trade (USD)49.7563.5578.71

Share of resampled paths ending at or below zero net profit: 0.0 %.

Streaks and recovery (block-bootstrap paths):

MetricMedianP90P95P99
Max losing streak (days)5779
Time under water (days)180310374498

Ruin probability (floor only) per prop profile and sizing. The safe-sizing answer is the first column at or below 10 %:

ProfileFloor %1/11/1.51/21/31/41/61/8Safe at
iqcapital_classic6.0100.0100.0100.0100.0100.0100.0100.0none ≤ 10 %
ftmo_challenge10.0100.0100.0100.0100.0100.0100.099.9none ≤ 10 %
generic_6pct_trailing6.0100.0100.0100.0100.0100.0100.0100.0none ≤ 10 %

Cells are breach probabilities in percent at each sizing (1/2 = half the tested lots). Floor only: the full rule set is stricter, so the truly safe sizing is at most the bold one.

Challenge pass probability, ftmo_challenge (at the tested sizing; daily and floor checks based on: intraday-sampled equity vs. day anchor):

StageTargetP(pass)P(fail: floor)P(fail: daily)UndecidedMedian days to pass
phase110 %40.8 %0.8 %58.4 %0.0 %6
phase25 %51.0 %0.8 %48.2 %0.0 %4

Both phases passed (independent-resample approximation): 20.8 %.

Attempt economics: expected attempts to pass phase 1: 2.5; for a 90 % chance of at least one pass: 5 attempts; probability of 5 consecutive fails: 7.3 %. multiply attempts by your challenge fee for the expected cost to fund.

Does reducing risk raise the pass chance? Phase 1 at each sizing (targets stay fixed, trading scales down):

Sizing1/11/1.51/21/31/41/61/8
P(pass)40.8 %45.8 %51.6 %55.5 %61.1 %60.2 %65.8 %
Median days681116192536
What these numbers can and cannot say:

Withdrawal replay and Capital what-if

Lot policy detected: mixed_or_unknown (lot-size CV 0.883, lots/balance CV 0.515). Read the matching row below.

Monthly withdrawal replay (monthly sweep to start balance), covering the active span 2020-01-02 … 2026-08-10. An EA can sit out most of the tested window, and all monthly and per-annum figures refer to this span:

Sizing modelMonths paidTotal withdrawnWithdrawn %/aMedian paid monthBest monthDry streak (months)
fixed_lots41/80271468397.272607324086
balance_scaled27/80267966392.152368997288

Monthly PnL heatmap: the dry-streak figure, visible at a glance:

JFMAMJJASOND20202020-01: +69+692020-02: -29-292020-03: -941-9412020-04: +1,396+1,3962020-05: +1,484+1,4842020-06: -2,361-2,3612020-07: +1,115+1,1152020-08: +916+9162020-09: +237+2372020-10: +94+942020-11: +1,566+1,5662020-12: -1,011-1,01120212021-01: -101-1012021-02: +1,071+1,0712021-03: -25-252021-04: -12-122021-05: -109-1092021-06: +1,120+1,1202021-07: -82-822021-08: +2,264+2,2642021-09: -1,662-1,6622021-10: +674+6742021-11: +504+5042021-12: +63+6320222022-01: -316-3162022-02: +1,199+1,1992022-03: +12+122022-04: -78-782022-05: +1,873+1,8732022-06: -1,613-1,6132022-07: +217+2172022-08: +2,304+2,3042022-09: -556-5562022-10: +104+1042022-11: -2,060-2,0602022-12: +174+17420232023-01: +2,870+2,8702023-02: -844-8442023-03: +2,411+2,4112023-04: +2,164+2,1642023-05: +175+1752023-06: -2,310-2,3102023-07: +450+4502023-08: -1,283-1,2832023-09: +3,704+3,7042023-10: +4,339+4,3392023-11: +2,457+2,4572023-12: +4,095+4,09520242024-01: +41+412024-02: +4,129+4,1292024-03: +5,995+5,9952024-04: +11,129+11,1292024-05: -1,738-1,7382024-06: -663-6632024-07: +11,687+11,6872024-08: -8,495-8,4952024-09: +3,995+3,9952024-10: +4,451+4,4512024-11: +6,100+6,1002024-12: +2,607+2,60720252025-01: +920+9202025-02: -3,520-3,5202025-03: +5,874+5,8742025-04: +12,148+12,1482025-05: +16,485+16,4852025-06: +11,657+11,6572025-07: -267-2672025-08: -1,388-1,3882025-09: +12,475+12,4752025-10: -11,006-11,0062025-11: +15,353+15,3532025-12: +21,720+21,72020262026-01: +32,408+32,4082026-02: +4,304+4,3042026-03: +20,345+20,3452026-04: +4,722+4,7222026-05: +18,984+18,9842026-06: +21,043+21,0432026-07: -6,272-6,2722026-08: +30,516+30,516monthly PnL (USD) of the observed backtest · grey = no trading days in the active span

Across 1000 resampled paths (fixed lots), total withdrawn spans 178803 … 270870 … 371624 USD (P5/median/P95); share of paths paying nothing at all: 0.0 %.

Capital what-if (fixed lots: identical trades, different account):

CapitalMaxDD %P(breach) iqcapital_classicP(breach) ftmo_challengeP(breach) generic_6pct_trailing
10001613.99100.0 % · observed hit100.0 % · observed hit100.0 % · observed hit
2500645.6100.0 % · observed hit100.0 % · observed hit100.0 % · observed hit
5000322.8100.0 % · observed hit100.0 % · observed hit100.0 % · observed hit
10000 (tested)161.4100.0 % · observed hit100.0 % · observed hit100.0 % · observed hit
2500064.56100.0 % · observed hit100.0 % · observed hit100.0 % · observed hit
10000016.14100.0 % · observed hit98.7 % · observed hit100.0 % · observed hit

fixed-lots USD path re-expressed per capital; margin not modeled. Safe sizing at another capital follows the MC ruin matrix scaled by capital/tested.

Assumptions and their direction:

Glossary

Pairing confidenceHow entry and exit deals were matched into trades: 'exact' = taken from the tester log; 'validated' = FIFO/LIFO reproduced the report's holding-time figures; 'heuristic' = unconfirmed FIFO assumption.
EOD Sharpe (annualized)Sharpe ratio computed from end-of-day balance returns, annualized with √252. Comparable across systems, unlike the report-head 'Sharpe', which is per-trade.
Cost share of grossCommission plus swap as a share of gross profit before costs. High values mean the edge is eaten by fees and financing.
Max drawdown (EOD)Largest peak-to-trough loss of the end-of-day balance curve.
Floating drawdown (sampled)Largest drawdown of reconstructed equity (balance plus open-position value), sampled at deal timestamps. Between deals equity is unobservable from a report; the tester-head equity DD is tick-based.
MartingalePosition sizing that grows after losses. Looks smooth for months, then loses the account in one streak.
Consistency ruleProp-firm rule capping the best day's share of total profit; punishes concentrated profit profiles.
Account deathsNumber of times the simulated account breached the profile's drawdown floor over this history (account is then reset and the simulation continues).
Withdrawal %/aYearly withdrawal as percent of account size in the sweep simulation (profits above start are swept daily).
Z-ScoreSerial correlation of the win/loss sequence. Strongly negative values often just reflect several sub-strategies interleaving, not necessarily a defect.
Monte-Carlo resamplingRe-arranging or re-drawing the audited trades many times to see the range of drawdowns and streaks the same trading could have produced. It cannot add information; it reveals path fragility, not future returns.
Block bootstrapBootstrap that draws whole multi-day blocks of the daily PnL series instead of single trades, preserving short-range clustering (losing weeks stay losing weeks).
Ruin probability (floor only)Share of resampled paths that breach the profile's drawdown floor at least once, with profits above start swept. The firm's full rule set is stricter, so this is a lower bound.
Time under waterLongest stretch of trading days a path spends below its previous balance peak.
Withdrawal replayRe-plays the backtest with a monthly payout: on the last trading day of each month, everything above the start balance is withdrawn. Shows what the strategy pays a trader who lives off it, instead of compounding like a backtest.
Dry streakLongest run of consecutive months in which the monthly withdrawal was zero: months a trader living off the account would have earned nothing.
Lot policyWhether the EA trades fixed lot sizes or scales them with the balance, detected from the variation of lot sizes across the deal list.
Capital what-ifThe observed USD path re-expressed against a different account size (fixed lots): drawdown percentages and floor-breach risk change with capital even though the trades are identical. Margin limits are not modeled.
Underlying vs. strategyThe symbol's own price move over the traded span, approximated from entry prices. If most profit is long-side while the symbol itself rose strongly, part of the result is the market's tailwind, not the mechanics. The tester cannot separate the two.
Cost fragilityHow quickly the result dies as per-trade costs rise. The break-even shock is the extra cost per trade (spread, slippage, commission) at which net profit reaches zero; high-frequency systems often die at cents.

What this audit cannot tell you

A good backtest cannot prove an edge. This audit dissects the simulation you gave it. It can expose structural risks (martingale, grids, hidden concentration, cost fragility, rule conflicts), but it cannot tell you that the strategy will make money. Three limits are fundamental:

This audit is software analysis, not investment advice, and contains no recommendation to buy or trade anything.