Twenty-three cents decides how this Mistral vs Claude for trading comparison ends, so start there: Claude's Season 6 head-to-head win — earned by holding near flat while Mistral fell — closed with a settled book of exactly +$0.23 once the season-ending liquidation swept every position. The record spans Seasons 5–6, the only completed TradeRank seasons the two share, since Mistral joined in Season 5; Mistral Medium 3.5 ran both, while the Anthropic slot ran Claude Opus 4.7 and then a Season 6 whose build handover the archive cannot place — which is why the pack credits that season to the slot and withholds its per-family behavior metrics. Both slots traded the same crypto list from a simulated $10,000 under one rulebook; every number here is pulled from the generated evidence pack linked at the end.
Which versions traded each season
| Season | Dates | Mistral version | Claude version | Asset universe | Field |
|---|---|---|---|---|---|
| Season 5 | May–Jun 2026 | Mistral Medium 3.5 | Claude Opus 4.7 | 10 crypto assets | 10 models |
| Season 6 | Jun–Jul 2026 | Mistral Medium 3.5 | Claude Opus 4.8 → Claude Fable 5 (boundary unprovable) | 10 crypto assets | 11 models |
Head-to-head results by season
| Season | Mistral return | Claude return | Gap (Mistral − Claude, pts, comparability basis) | Rank (Mistral / Claude) | Trades (Mistral / Claude) | Win rate (Mistral / Claude) | Max drawdown (Mistral / Claude) | Winner |
|---|---|---|---|---|---|---|---|---|
| Season 5 | +9.55% | +2.67% | +6.88 | 3rd of 10 / 6th of 10 | 6 / 13 | 83.3% / 69.2% | 7.94% / 11.69% | Mistral |
| Season 6 | -2.95% | -0.23% | -2.73 | 9th of 11 / 5th of 11 | 8 / — (excluded) | 25% / — (excluded) | 6.24% / — (excluded) | Claude |
Returns, season by season

How the Mistral vs Claude for Trading Series Split
Neither win requires much narrative. In Season 5 both models made money and Mistral made distinctly more — +9.55% to +2.67%, a +6.88-point gap, 3rd in the field to Claude's 6th. In Season 6 Claude's -0.23% was close enough to flat to place 5th of 11, and Mistral's -2.95% was not, landing 9th. A wide win in a green season, a narrow win in a red one, one apiece.
What gives the series its texture is how little the winning had to do with banking money. Mistral's Season 5 win rode mostly on open marks; Claude's Season 6 win settled for +$0.23. The one truly solid settled profit either account produced — Claude's +$324.41 in Season 5 — belonged to a losing season. Realized P&L is a settlement fact, not an alternative scoreboard, and this pair is a 2-season course in why the distinction matters.
Each season's return against its deepest drawdown

A Win Settled for a Quarter of a Dollar
The realized column of this pair holds its strangest cell. Claude's Season 6 — the season it won — closed with a settled book of exactly +$0.23: a month of orders, force-liquidated at the close, netting out to almost nothing. The win was real on the standings, and on banked cash it was a wash. Mistral's losing side of that season settled at -$276.24, with a -$18.85 residue in the standings' unrealized column, so on that side the season's verdict and its cash verdict at least shared a sign.
Season 5 was the more conventional split, with an ironic edge. Claude banked +$324.41 — the pair's largest settled profit — while holding a small open loss of -$57.40, and lost the season; Mistral won it on a book of +$180.39 banked under +$774.47 of marks. Across the 2 seasons, banking the most money and winning the season described different accounts in Season 5 — and in Season 6 the winner did out-bank the loser, while banking only $0.23 itself.
Trade count by season

What the Earliest Attributable Decisions Recorded
Season 5's cycles are where the archive can attribute individual outcomes, and the selection rule is mechanical: each slot's earliest attributable gain and earliest attributable loss, read from position-state changes between consecutive daily snapshots. Claude's (Claude Opus 4.7) earliest gain was a SUI short — its log reads "Clean trend-follow short." — and its earliest loss an ETH short from the season's opening cycle that the next snapshot marked underwater. Mistral's earliest gain was a TRX long on that same opening cycle, the one ticker to clear its entry rules that day; its earliest attributable loss, a DOGE short, is dated mid-June.
These are reconstructions, not fills — a position opened and closed within a cycle leaves nothing to reconstruct — and the pack stores no sizes or exits around them. What they preserve is small and specific: Claude's first recorded outcomes came from short positions, Mistral's first from the lone long its rules admitted, and each slot logged one early gain and one early loss from those entries.
How We Measured the Pair
The Season 6 label problem comes first, because this pair wears it twice over: the Anthropic slot entered Season 6 as Claude Opus 4.8 and finished it as Claude Fable 5, the archive cannot place the handover, and the pack responds by crediting the season to the Claude slot and withholding its per-family Season 6 behavior metrics — trade count, win rate, maximum drawdown — from comparison. Only Season 5 supports version-level or behavior-level statements about Claude here, and this page makes none beyond it. Mistral Medium 3.5, by contrast, ran both seasons unchanged.
Everything else is the standard rig. Identical per-season conditions for both slots — the same 10-asset list, one decision a day, $10,000 simulated, live prices, a modeled 0.1% fee, no slippage or borrow costs — with the field growing from 10 to 11 models between seasons, the prompt regime switching mid-Season 6 to a medium-term investor framing, and — per the era notes — a universe expansion toward US equities landing mid-season at a point the archive cannot fix. Season 6 ended in a forced liquidation, so the pack computes its gap from pre-liquidation snapshots (Mistral -3.02%, Claude -0.30%, the -2.73 in the table) while the return columns print the official standings. A deterministic program gathers every figure from the archived reports, decision logs and equity snapshots into the evidence pack, and this page must reproduce the pack — its content hash is the arbiter — before it publishes.
Limitations: Two Observations, One Unprovable Label
The sample is the first wall: 2 shared completed seasons is 2 observations, and a 1-1 on 2 observations is the least conclusive record this format can produce. The second wall is attribution: Claude's Season 6 behavior is withheld by the pack, so any comparison of how the 2 slots traded — activity, risk, win rate — rests on Season 5 alone, where Claude ran the busier book (13 orders to 6), the deeper drawdown (11.69% to 7.94%) and the lower win rate (69.2% to 83.3%, both report figures that count open positions). The $0.23 cell, memorable as it is, is a settlement fact about one force-liquidated season, not a verdict on how Claude traded it. Season 5's returns lean on open marks for Mistral especially, and 'Claude beat Mistral in Season 6' is exactly as specific as the archive allows — the slot won; which build won is unknowable. The live LLM trading benchmark will add the third observation when the next shared season closes; until then, every figure on this page can be checked in the Mistral vs Claude evidence pack.