30-Day BacktestSwing Trader OS · Phase 1
Day
Cases
0
Done
0/—
Streak
0
Phase 1 · Weeks 1–4 · 30 Days

Manufacture evidence, not opinion.

The backtest phase produces the dataset the rest of the system runs on. 35+ logged cases. Win Rate computed. R:R computed. Personal Playbook built from your data — not someone else's setups. One Gate Check at the end. Pass or repeat.
35+ cases · 4 review days · ~2 hrs/day · 2-hour fallback
§A
Daily Operating Protocol
~2h standard · 2h fallback also valid
STANDARD
~2 hours, in order
120 min
DRC
§1 Temperature Check · sets Risk Mode (5–10 min)
DRC
§2 Market Context · BTC bias · BTC.D · ETH/BTC · regime · 4H/1H structure · narrative (10–15 min)
ATLAS
Reference 7-condition checklist (after Day 24, Atlas is the source) (10 min)
BT
Open Replay → Pre-Trade case → advance bars → Post-Trade + Execution + Outcome (60–90 min)
DRC
§5 EOD Process + Adherence Score · §6 Self-Regulation Audit (10–15 min)
FALLBACK
2-hour cap (tight days)
120 min
DRC
§1 Temperature (5 min)
DRC
§2 Market Context — short form (10 min)
BT
Core backtest — minimum 1 case, ideally 2 (90 min)
DRC
EOD + TCJ close (15 min)
Fallback is a deliberate downgrade, not a skip. The DRC streak must remain intact across all 30 days.
§B
What to Record · Where
DRC · TCJ · Weekly Review · Atlas
EventDRCTCJWRAtlas
Morning startD§1 Temp → Risk ModeRead-only (Day 25+)
Market readD§2 ContextBias + Regime check
Open BT caseT§0 + §1 Pre-TradeRead §3 checklist
Close BT caseT§2 + §3 + §4Queues weekly signal
No trade todayD§3 → No-TradeTNo-Trade Case
End of sessionD§5 + §6 EOD
Loss patternA§7 Error Library
Review day (7,14,21,30)§1 + §5 + §6Sync ReportW§1–§8 full§6 Best/Worst · §7 reconcile
W1
Week 1 · Setup & Learning
Days 1–7 · Environment + Replay + Example cases
W2
Week 2 · 20 BTC Backtest Trades
Days 8–14 · Log only · no judgment
W3
Week 3 · ETH + Advanced Conditions
Days 15–21 · ETH + BTC.D + false-setup detection
W4
Week 4 · Playbook + Readiness
Days 22–30 · Atlas build · Drawdown · Gate Check #1
§C
Valid Backtest Case · 10 Criteria
All must be true · else Case Incomplete
1
Case Type = Backtest in TCJ §0 · Phase = 1
2
Ticker = BTC/USDT or ETH/USDT only — no other pairs in Phase 1
3
Pre-Trade written before outcome bar visible in Replay — no result contamination
4
Thesis contains structure ref + trigger ref + driver ref (all three)
5
Invalidation is price-precise — a number or "1H close below X" — not a feeling
6
R:R ≥ 1.0 at planned TP1
7
Post-Trade closed same session — no overnight "I'll finish tomorrow"
8
Sections 2 + 3 + 4 all filled at close
9
Outcome computed from R Realized, not assigned manually
10
Lesson uses process language, not P&L language
§D
Invalid Cases · Recognized Failure Modes
Rejected from rollups · flagged for repair
FailureWhy It's Rejected
Thesis without triggerCan't separate setup from prediction
Invalidation as feelingNo measurable failure point
Outcome bar visible at Pre-TradeResult-knowledge contamination
Grade C taken as normal tradeViolates DRC §3 block · backtest-study only
Edited Thesis after entryTCJ §3 #12 self-deception breach
Lesson written in P&L languageWrong layer — process, not P&L
No-Trade with no categoryAvoidance, not discipline
Pair outside Atlas scopeSetup outside system (only BTC, ETH)
Timeframe below 1H referencedClass A1 ban (T3)
Skipped §2 / §3 / §4 at closeCase Incomplete
§E
Labeling · Wins · Losses · No-Trades · Breaches
Outcomes by R Realized · Plan Match orthogonal
LabelDefinitionCounts In
WinR Realized ≥ +0.8R · all sections completeWin Rate numerator
LossR Realized ≤ −0.9R · all sections completeWin Rate denominator
BreakevenR Realized −0.2R to +0.2RExcluded from Win Rate
Scratch−0.9 to −0.2 or +0.2 to +0.8 (manual exits)Flagged in Execution Score
No-TradeSetup not taken · valid reason categoryDiscipline Counter
Rule BreachAny of 12 TCJ §3 breach itemsProcess Score component d
Case IncompleteFailed §C validity rulesExcluded · repair in WR
Plan Match is orthogonal to outcome.
A Loss with Plan Match = Full and Was Thesis Correct? = Y is a clean loss — it strengthens the playbook, not weakens it. A Win with Plan Match = Broken is a sloppy win — it does not strengthen the setup; it surfaces a behavioral risk for the Error Library.
§F
Worked Examples
3 good · 5 bad · all on BTC / ETH

§G · TCJ Sync Report

Run weekly · Days 7, 14, 21, 30
Win Rate
Avg R Realized
Avg R:R Planned
Expectancy (per trade, R)
Set Day # above 0 to assess against the case-count floor.
§H · Gate Check #1 — Day 30
Tap the dot to mark each criterion green. Five of these five must be green to enter Phase 2. Criterion #5 activates in Phase 3 only.
#1
Win Rate ≥ 45%
across 35+ Backtest cases · TCJ Sync Report
#2
Avg R:R ≥ 1.2
across closed cases · TCJ Sync Report
#3
Atlas Playbook Completeness ≥ 70%
§1, §2, §3, §5 complete for all three setups
#4
Drawdown Protocol Written
Atlas §0 toggle = Y · non-empty §1 text · 3-loss rule locked
#6
DRC Streak 30 / 30
no skipped days, even No-Trade days
Gate #1 Result
Mark criteria to compute decision
§I
How to Think During Backtesting
8 frames · execution language only
1
Each case is a sentence the system can read. Not a story you tell yourself. If TCJ would reject the case, the trade didn't happen.
2
Result is downstream of process. A Win that broke plan is information about you, not the setup. A Loss that respected invalidation is information about the setup, not you.
3
The bar at Pre-Trade is the only state you're allowed to know. Replay's job is to enforce this. Yours is not to cheat it.
4
No-Trade is a valid output. A day with one A-grade No-Trade is stronger than a day with three impulse cases.
5
Numbers come from quantity, not certainty. 35 cases produce a Win Rate. 8 cases produce a feeling. Hit the case floors before drawing conclusions.
6
The Error Library is the most valuable file in the system. Wins teach a setup; losses teach a setup; broken-plan trades teach you yourself.
7
Atlas comes from data, not opinion. Days 22–24 codify the patterns Days 8–21 already proved. Atlas does not invent.
8
A Gate Check #1 fail is information, not failure. The 14-day extension exists because entering Phase 2 with bad data is worse than a delay.