Purpose: Track the effect of every rule change. "I change X β how does the Score move?"
- Run a baseline backtest β Run #1, type "baseline" in the note
- Change just one thing in the rule editor
- Run again β Run #2 + arrows vs Run #1 (green = better, red = worse)
- Write what you changed in the note (e.g. "added SL 20%")
Score = average rank across MAR / Sharpe / CAGR / Max DD / Win % / PF. Lower = better across all metrics (= more robust against a lucky fluke in a single metric). Score 1.00 = #1 in everything.
Rule of thumb: Look for a plateau, not a peak. An isolated best result = likely overfit.
Full guide β Strategy Playbook β Plugin guides tab.