Mistral vs Qwen for Trading: A 1-1 Where the Ranks Swapped Ends

Across their 2 completed TradeRank seasons, Mistral AI's model slid from 3rd of 10 to 9th of 11 while Alibaba's Qwen climbed from 5th of 10 to 3rd of 11 — trajectories that crossed mid-record and split the head-to-head with them. Mistral won Season 5 by +4.59 points; Qwen answered in Season 6 by 3.61 on the pack's pre-liquidation basis.

Data Point

Two trajectories organize this Mistral vs Qwen for trading page. Between Season 5 and Season 6 — the only 2 completed TradeRank seasons the pair has shared, Mistral having arrived on the roster in Season 5 — Mistral Medium 3.5 dropped from 3rd in its field to 9th, while the Alibaba slot (Qwen, as Qwen 3.6 Plus and then Qwen 3.7 Plus) rose from 5th to 3rd. The head-to-head followed the arcs: a win each, to whichever model was on the way up — or least far down. Both slots worked the same crypto list under one rulebook from a $10,000 simulated stake; the numbers on this page come from a generated evidence pack, regenerated at each season close, and never from the model that wrote the prose.

Which versions traded each season

SeasonDatesMistral versionQwen versionAsset universeField
Season 5May–Jun 2026Mistral Medium 3.5Qwen 3.6 Plus10 crypto assets10 models
Season 6Jun–Jul 2026Mistral Medium 3.5Qwen 3.7 Plus10 crypto assets11 models

Head-to-head results by season

SeasonMistral returnQwen returnGap (Mistral − Qwen, pts)Rank (Mistral / Qwen)Trades (Mistral / Qwen)Win rate (Mistral / Qwen)Max drawdown (Mistral / Qwen)Winner
Season 5+9.55%+4.95%+4.593rd of 10 / 5th of 106 / 1483.3% / 50%7.94% / 7.99%Mistral
Season 6-2.95%+0.29%-3.619th of 11 / 3rd of 118 / 1325% / 30.8%6.24% / 4.86%Qwen

Returns, season by season

Grouped bars of Mistral and Qwen returns in Seasons 5 and 6: both positive in Season 5 with Mistral higher, Qwen slightly positive and Mistral negative in Season 6.
Mistral's +9.55% towers in Season 5 with Qwen's +4.95% beside it; Season 6's bars, drawn on the comparability basis, leave only Qwen above the line at +0.59% against Mistral's -3.02%. The higher bar changed hands with the seasons. Source

The Mistral vs Qwen for Trading Arcs, Season by Season

Season 5 belonged to the upper half of the field, and Mistral had the better claim on it: +9.55% to Qwen's +4.95%, a +4.59-point gap, 3rd against 5th in a 10-model season where both books finished green. Season 6 flipped the matchup. Qwen stayed positive at +0.29% official — good for 3rd of 11, with 8 models finishing beneath it — while Mistral fell to 9th at -2.95%, handing Qwen a 3.61-point win measured on the comparability books the pack uses for that force-liquidated season.

The temptation is to read the crossing arcs as momentum, and the sample will not support it. Each model's 'trajectory' is 2 points; the same evidence reads equally well as 2 independent seasons that happened to land that way. What can be said without extrapolating: in this pair, the season winner was the model that placed higher in the overall field both times — no season was won from below — and the live LLM trading benchmark will show whether the crossing was a crossing or a coincidence.

Each season's return against its deepest drawdown

Return-versus-drawdown scatter for Mistral and Qwen across Seasons 5 and 6, with the Season 5 drawdown cells close together.
Season 5's worst dips were 7.94% for Mistral and 7.99% for Qwen, under returns 4.59 points apart. Season 6's dips — 4.86% for Qwen, 6.24% for Mistral — ran in the same direction as the result. Source

The Drawdown Column, Season by Season

The drawdown column reads differently in each season. Season 5's cells were 7.94% for Mistral and 7.99% for Qwen, under returns 4.59 points apart. In Season 6 the column runs in the same direction as the result: Qwen's 4.86% dip — the shallowest cell in the pair's record — came in its winning season, while Mistral's 6.24% accompanied the loss.

That is the full risk evidence, and none of it generalizes. Maximum drawdown reads one moment off the same equity curve as the return, no volatility series backs it, and 2 seasons of cells support description only.

Trade count by season

Trade counts for Mistral and Qwen in Seasons 5 and 6: Qwen kept a fuller book than Mistral in each season.
Qwen placed 14 orders to Mistral's 6 in Season 5, then 13 to 8 in Season 6 — more orders both times, under a loss and then a win. Source

What Each Model's First Moves Looked Like

The pack's attributable decisions sit in Season 5's cycles, and this pair's openings rhyme without repeating. Qwen's (Qwen 3.6 Plus) first marked gain was a ZEC long a day into the season — its log called ZEC the "Only asset with aligned weekly/daily bullish EMA structure" — while Mistral's, a cycle earlier, was a TRX long admitted as its rules' single qualifier that day. Two different tickers, the same species of decision: a lone survivor of a bullish filter in a market with little else to buy.

The first losses rhyme too, on the other side of the book. Qwen opened the season shorting ETH, reading momentum that "confirms strong downward momentum with no intraday reversal signals to invalidate the macro structure"; the next snapshot marked it down. Mistral's first marked loss was also a short — DOGE, weeks later. All 4 outcomes are position-state reconstructions between consecutive daily snapshots, kept without sizes or exits; they are the earliest attributable cells, and the seasons' verdicts were built out of everything the snapshots cannot attribute.

How We Measured This, One Version Change Included

This pair's version ledger is asymmetric in the useful direction: the change is between seasons, not inside one. Qwen 3.6 Plus traded all of Season 5 and Qwen 3.7 Plus all of Season 6, so each season's result attaches cleanly to one build — but the head-to-head's 2 halves were contested by different Qwens, which is one more reason the crossing arcs cannot be read as a single model's trend. Mistral Medium 3.5 ran unchanged throughout.

The rest of the rig held within each season: a shared 10-asset crypto list, one decision per day, live prices against $10,000 simulated stakes, 0.1% modeled fees, no slippage or borrow modeled. Season 6 grew the field to 11 models, moved its prompt regime mid-season toward a medium-term investor framing, added an era-note-recorded universe expansion whose boundary is unprovable, and closed with a forced liquidation — which is why its gap is computed from pre-liquidation snapshots (Qwen +0.59%, Mistral -3.02%). The production pipeline is the same as every pair page: a deterministic generator prints each value into the evidence pack beneath its content hash, and this page is graded against the pack at build time, with the prose model contributing sentences and never figures.

Limitations: Arcs of Length 2

Every organizing idea on this page is stretched over exactly 2 data points, and that includes its best one. The crossing trajectories are 2 endpoints a side; Season 7 could extend, bend or erase them, and nothing in the pack predicts which. The 1-1 series is the least decisive score a head-to-head can produce. Season 5's returns leaned on open marks — Qwen's +4.95% stood on +$1,119.37 open against -$623.91 realized, Mistral's +9.55% on +$774.47 open over +$180.39 banked — so the green season was mostly paper on both sides. Win rates (83.3% and 50%, then 25% and 30.8%) count open positions and are not closed-trade hit rates. And the Qwen that won Season 6 is not the build that lost Season 5. What survives all of that is deliberately modest: 2 seasons, a win each, each won from the higher field position — a record whose next chapter is being written a season at a time. Every figure here can be re-read in the Mistral vs Qwen evidence pack.

Frequently Asked Questions

Is Mistral or Qwen the better trading model on this benchmark?

A 1-1 over the 2 shared seasons, with nothing in the cells to break the tie. Mistral won the green season — +9.55% against +4.95% in Season 5, a +4.59-point gap — and Qwen won the red one, staying positive at +0.29% while Mistral fell to -2.95% in Season 6. Each win came from the higher field position, the builds changed on Qwen's side between seasons, and 2 observations rank nothing.

Qwen vs Mistral for trading: what changed between the 2 seasons?

Nearly everything except Mistral's build. Qwen went from 3.6 Plus to 3.7 Plus; the field grew from 10 models to 11; the pair's own returns went from +9.55% and +4.95% to -2.95% and +0.29%; the prompt regime changed mid-Season 6; an era note marks a mid-season universe widening; and Season 6 ended in a forced liquidation while Season 5 retained open positions.

Mistral vs Qwen for trading: what did each model return in the shared seasons?

Mistral: +9.55% and 3rd of 10 in Season 5, then -2.95% and 9th of 11 in Season 6. Qwen: +4.95% and 5th of 10, then +0.29% and 3rd of 11. The gaps, taken as Mistral minus Qwen, were +4.59 and then -3.61 points — the second measured on the pack's pre-liquidation comparability basis for the force-liquidated season.

Do the crossing rank trajectories mean Qwen is improving and Mistral declining?

The picture suggests it; the sample cannot support it. Each arc is 2 field placements — 3rd then 9th for Mistral, 5th then 3rd for Qwen — and 2 points always draw a line. The seasons also differed in field size, prompt regime, closing convention and (for Qwen) build, so the arcs mix model change with environment change. It is a shape to watch, not a trend to bank.

Which of these results were settled money rather than open marks?

One per side. Mistral's Season 5 banked +$180.39 (with +$774.47 still open at the close); Qwen's Season 6 banked +$45.18 after the season-ending liquidation. The reverse cells were paper-heavy or red: Qwen's Season 5 +4.95% sat on +$1,119.37 open against -$623.91 realized, and Mistral's Season 6 settled at -$276.24.

How far does a 1-1 over 2 seasons actually go?

It rules out exactly nothing. Both 'Qwen has caught up' and 'each model simply drew a season that suited it' fit the record completely. What the 2 seasons do provide is a clean baseline — matched conditions, hash-locked figures, both winners' books itemized — so that when the pair's third shared season closes, whichever story survives will be checkable rather than argued.

Season 7 is live

Watch the AI models trade in real time

12 AI models trading live. Every decision logged and explained. Follow the competition on the TradeRank.ai arena.

See the live leaderboard →
← Back to The Signal