Research Register

Making a backtest look strong is the easiest exercise in trading.

The rules we measured, and what held.

You start at the edge of what has already been tested, instead of spending months on a direction that closed here in a week.

  • Free The question and the scope of every entry.
  • VIP The measurement, the verdict and its boundary.

This page and the MCP tool get_research_verdicts read the same table, so your assistant answers with exactly what stands here.

14 entries in two baskets.  State of 12 August 2026.

Basket 1

Ideas traders bring to us

6 entries

Candle strength and level confluence as a mechanical gold entry (XAUUSD M5)

Showcase

Both readings point the same way on gold. If I take only the M5 entries where a candle strength signal lines up with a support or resistance level, does the confluence lift my hit rate?

  • Instrument Gold (XAUUSD)
  • Timeframe M5
  • Window January to May 2026

One entry stays open for everyone. This is what a full one looks like.

What we measured
XAUUSD M5, January to May 2026, 28,861 bars. Each candidate entry was taken as a pending stop order with a fixed target at twice the risk, costs included. First the base rule alone, then each confluence condition added on top of the identical decision points, then the full stack, so every filter was judged against the same base population.
Verdict class
Falsified
Verdict
Falsified. The base rule wins about 26% of its trades at a 2:1 target, and it needs 33.3% to break even. None of the confluence conditions lifted the win rate across that line. The full stack fired 22 trades in three months, a sample too thin to carry any decision.
What this means for you
On this instrument, this timeframe and this window, agreement between the two readings failed to select better entries. Both remain what they are, a way of reading structure and context by eye, and the entry decision that follows stays a discretionary one. Our measurement offers you no mechanical shortcut for it.
What stays open
One instrument, one timeframe, one target logic, five months. Other targets, other exit logic, other instruments and a workflow with a human in the loop were outside the test. The last of these cannot be measured this way at all.

Currency strength breadth as a mechanical signal (28 FX pairs)

When one currency is strong across the board and another is weak across the board, that should be the cleanest trend there is. Can I trade that spread mechanically, and if the extremes revert, can I fade them?

  • Instrument 28 FX pairs
  • Window January to May 2026
VIP

What we measured, the verdict, what it means for your own testing and what stays open: VIP opens all four.

A grid as an edge on an index (US500)

Does the grid logic from a gold system carry over to an index like US500?

  • Instrument Index (US500)
  • Window 2018 to 2026
VIP

What we measured, the verdict, what it means for your own testing and what stays open: VIP opens all four.

A grid portfolio on range bound FX crosses

Range bound crosses like AUDCAD, AUDNZD and NZDCAD look made for mean reversion. Does a basket of them diversify a grid?

  • Instrument FX crosses (AUDCAD, AUDNZD, NZDCAD)
  • Window 4.5 years
VIP

What we measured, the verdict, what it means for your own testing and what stays open: VIP opens all four.

Futures term structure and traded volume as timing signals (gold, FX, index futures)

CFDs hide the real market. If I read the futures curve and exchange volume, do I get a timing edge my CFD chart cannot show me?

  • Instrument Futures (gold, FX, index)
  • Timeframe H1
  • Window 3,101 trading days
VIP

What we measured, the verdict, what it means for your own testing and what stays open: VIP opens all four.

Comparing tick volume between a CFD broker and an exchange feed

My broker prints tick volume. If I compare it against exchange volume, can I see where the real flow is?

  • Instrument Futures and CFD feeds
VIP

What we measured, the verdict, what it means for your own testing and what stays open: VIP opens all four.

Basket 2

Refinements we tested on our own systems

Every entry here started as a plausible improvement to a shipped system. Each was measured against the shipped version on identical data with identical costs. The pattern in this basket is the point: most improvements that sound obvious cost money once they are measured.

8 entries

A gold breakout approach: entering at the breakout instead of waiting for confirmation

The system misses vertical moves because it waits for confirmation. Why does it skip the market entry right at the breakout, or a hybrid of both?

  • Instrument Gold (XAUUSD)
VIP

What we measured, the verdict, what it means for your own testing and what stays open: VIP opens all four.

A gold breakout approach: a structure based initial stop instead of the built in distance

A swing pivot tool finds the real structural pivot. Why does the system place its initial stop somewhere else?

  • Instrument Gold (XAUUSD)
VIP

What we measured, the verdict, what it means for your own testing and what stays open: VIP opens all four.

A candle strength indicator: the history window as a stability or quality dial

The indicator lets me choose how much history it reads. Is a longer history the safer setting, and does agreement between two history lengths mark the better signals?

  • Instrument 10 symbols
VIP

What we measured, the verdict, what it means for your own testing and what stays open: VIP opens all four.

A gold basket approach: a daily trailing stop on basket profit

The basket gives profit back during the day. Should it lock in a share of the day's gain and stop for the day?

  • Instrument Gold (XAUUSD)
  • Timeframe H1
  • Window 1 January 2025 to 1 August 2026
VIP

What we measured, the verdict, what it means for your own testing and what stays open: VIP opens all four.

A gold basket approach: identifying the losing trades before entry

If the give back is real, the losing trades should look different at entry. Can a marker sort them out in advance?

  • Instrument Gold (XAUUSD)
  • Timeframe H1
VIP

What we measured, the verdict, what it means for your own testing and what stays open: VIP opens all four.

A candle strength trading system: which cohort lost the money in live trading

The live account trails the backtest. Where does the loss actually sit?

  • Instrument FX (live account, AUD and NZD basis)
  • Window April and May 2026
VIP

What we measured, the verdict, what it means for your own testing and what stays open: VIP opens all four.

A news basket approach: weighting risk toward the strongest leg of an event basket

When a release hits, some pairs move much harder than others in the first minutes. Should the basket put more risk on those legs?

  • Instrument FX (news event baskets)
  • Window August 2025 to July 2026
VIP

What we measured, the verdict, what it means for your own testing and what stays open: VIP opens all four.

A news basket approach: handling a counter impulse on an open basket

A second release turns the currency around while the first basket is still open. Should the old basket be closed, or at least its weakest legs cut?

  • Instrument FX (news event baskets)
  • Window Twelve months
VIP

What we measured, the verdict, what it means for your own testing and what stays open: VIP opens all four.

The standard

How we measure

Every verdict here was produced the same way. A Python simulation only sorts out what is hopeless. The word tradeable is earned in the MT5 Strategy Tester on real ticks, with the spread, commission and swap of a live symbol. The rule is fitted on one half of the data and applied unchanged to the other. Sample size, a bootstrap interval, an outlier check and a stability check all have to pass before a result counts.

The nine failure modes we check first

  • F1 Same bar resolution bias

    Stop and target inside the same bar, where the simulation has to guess which came first.

  • F2 Forward leak in indicator timing

    A value that was still forming at the moment of the decision.

  • F3 Cost omission

    Results computed before spread, commission and swap.

  • F4 Truncated lookahead

    A fixed forward window that silently drops every trade resolving later.

  • F5 Missing robustness battery

    One sample, no bootstrap interval, no out of sample half, no stability check.

  • F6 Snapshot trap in basket simulations

    The best excursion seen so far, read back as the price at that moment.

  • F7 Tester cache and history quality

    Reused inputs from the last run, and a run on thin history.

  • F8 Single point optimum

    One bright cell whose neighbours are poor. A finding is a plateau, a spike is noise.

  • F9 Cost drift between runs

    Two runs of the same system carrying very different cost per trade.

Where a study searches through many candidates, the hypothesis, the cells and the kill criteria are written down before the first number is computed, and the best candidate is compared against the best of an equally large set of random ones. All of this establishes what held on the tested instrument, in the tested window, under real trading costs. Live results decide the rest, and they are the last stage of every cascade we run.

Ask your assistant

The same register, inside your chat

Point your MCP client at our endpoint and call get_research_verdicts. Ask by topic, by symbol or by entry, and you get what stands on this page, with the same cut: the question and the scope for everyone, the verdict for VIP keys.

https://stein.investments/mcp

Tool: get_research_verdicts  ·  How to connect →

Notes

What the figures are

All figures are results measured on the data and in the windows named in each entry. They are past observations and carry no forecast and no recommendation to trade. Trading carries risk of loss. Parameter values, thresholds, filter logic and engine internals of our products stay with us. What we publish here is the question, the method, the result and its boundary.