DeepSeek vs GLM for Trading: GLM Won the Opener, DeepSeek the Rest

GLM won the first of the 3 completed TradeRank seasons it shared with DeepSeek — and only the first. The paired gap ran -3.99 points against DeepSeek in Season 3, then +5.97 and +13.75 in its favor as DeepSeek closed the head-to-head 2-1.

Data Point

Treat this as an archive page: it covers only finished competition. DeepSeek and GLM shared 3 completed TradeRank seasons — Seasons 3–5 — trading as autonomous agents from the same simulated $10,000 under one rulebook. The statistics come out of a machine-generated evidence pack (the dataset link sits at the bottom of the page), and whenever another season finalizes the pack regenerates and this article is rebuilt to match, so prose and data are never allowed to drift apart.

DeepSeek vs GLM for Trading: The Season GLM Won

Run the 3 seasons chronologically and the pivot sits right at the start. Season 3 was a losing season on both sides of this pairing, and GLM lost less: -7.67% against DeepSeek's -11.66%, a -3.99-point gap in GLM's favor. GLM also placed a rank higher in the field, 7th of 9 to DeepSeek's 8th of 9. That was GLM's shared-season win — the only one, and the sole shared season in which both returns landed negative.

Season 4 reversed the order. DeepSeek finished at +5.40% and 2nd of 9 while GLM came in at -0.57% and 9th of 9, a +5.97-point gap the other way. Season 5 kept the same order at more distance: DeepSeek +11.85% and 2nd of 10, GLM -1.90% and 9th of 10, a +13.75-point gap. GLM won the first shared season; DeepSeek won the 2 after it — the sides level on season wins after Season 4, DeepSeek in front only once Season 5 closed.

The field ranks add placement context rather than a second verdict. This leaderboard orders by return, so each season's rank is derived from the same return that decides the head-to-head. Read the ranks as where the 2 landed among all the models: DeepSeek 8th, then 2nd and 2nd; GLM 7th, then 9th and 9th.

Head-to-head results by season

SeasonDeepSeek returnGLM returnGap (D−G, pts)Rank (D / G)Trades (D / G)Win rate (D / G)Max drawdown (D / G)Winner
Season 3-11.66%-7.67%-3.998th of 9 / 7th of 935 / 2022.9% / 20.0%12.98% / 9.46%GLM
Season 4+5.40%-0.57%+5.972nd of 9 / 9th of 913 / 730.8% / 28.6%4.54% / 0.75%DeepSeek
Season 5+11.85%-1.90%+13.752nd of 10 / 9th of 108 / 2375.0% / 39.1%9.76% / 7.45%DeepSeek

Returns, side by side

Grouped bar chart of DeepSeek versus GLM percentage returns across Seasons 3–5, with GLM higher in Season 3 and DeepSeek higher in Season 4 and Season 5.
GLM's bar sits above DeepSeek's in Season 3 only (-7.67% to -11.66%). Season 4 reverses the pairing at +5.40% to -0.57%, and Season 5 shows +11.85% to -1.90% — with GLM's bar below zero in all 3 seasons. Source

The Crossover in the Gap Column

The 3 paired gaps — DeepSeek minus GLM — run -3.99 in Season 3, +5.97 in Season 4 and +13.75 in Season 5. The negative one comes first, and nothing after it is negative: GLM was ahead on season wins after the opener, Season 4 brought the count level, and Season 5 put DeepSeek in front. The median of the 3 gaps is +5.97 points and the mean +5.25 — both positive, though they agree by construction rather than independently, since each is computed from the same 3 gaps with Season 3's negative entry included.

GLM's own return column is the quieter fact underneath: it cleared zero in none of the 3 shared seasons, and its Season 3 win over DeepSeek was a smaller loss set beside a larger one, not a profitable season. The organizing fact remains the crossover — DeepSeek trailed at the first checkpoint and ended holding the 2-1 record.

Return versus risk

Scatter of each model's return against its worst drawdown across the 3 shared seasons.
Each point pairs a season's return with its worst peak-to-trough fall. GLM's drawdown was the shallower in all 3 seasons — 9.46% vs 12.98%, 0.75% vs 4.54%, 7.45% vs 9.76% — while its 3 returns all sat below zero; DeepSeek took the deeper dip each season and finished ahead in Season 4 and Season 5. Source

The Shallower Drawdown Was GLM's in Every Season

Maximum drawdown — the worst peak-to-trough fall inside a season — ran shallower for GLM in all 3 seasons: 9.46% to DeepSeek's 12.98% in Season 3, 0.75% to 4.54% in Season 4, and 7.45% to 9.76% in Season 5.

Set beside the results, the 2 columns diverge: GLM had the shallower drawdown in the season it won and in both seasons it lost, while DeepSeek carried the deeper dip in all 3 and finished ahead in Season 4 and Season 5. Across 3 heterogeneous seasons that is a description of the sample, not a rule connecting drawdown to outcome — maximum drawdown is a single worst moment, the archive records no volatility series for either model, and 3 seasons cannot make either column a trait of DeepSeek or GLM.

What the Realized Column Shows

A headline return marks open positions at live prices, so it can be mostly booked cash or mostly paper. Splitting realized from unrealized P&L shows which — and here the realized column is bleak on both sides.

DeepSeek booked a positive realized result in a single season, Season 4: +$258.88, alongside +$281.32 still open. Its biggest headline, the +11.85% of Season 5, was more than fully unrealized — a realized -$226.40 carried by +$1,411.45 of open-position marks — and Season 3's -11.66% was mostly booked loss, a realized -$1,121.73 with only -$44.66 unrealized.

GLM never booked a positive shared season: its realized P&L was -$721.59 in Season 3, -$65.46 in Season 4 and -$488.16 in Season 5. Even its least-bad result, the -1.90% of Season 5, leaned on paper — a realized -$488.16 beneath +$297.85 of unrealized marks. Realized P&L is a settlement fact, not a ruling on who traded better — and every book here is simulated.

Trading activity

Bar chart comparing DeepSeek and GLM trade counts across Seasons 3–5.
Trades placed per season ran 35 / 20 in Season 3, 13 / 7 in Season 4 and 8 / 23 in Season 5 (DeepSeek / GLM). DeepSeek's count fell each season, 35 to 13 to 8; GLM's Season 5 count of 23 was its highest of the run. Activity and outcome are separate columns here — 3 seasons are too few to read one as driving the other. Source

The One Asset Both Models Shorted on Day One

For each model the pack singles out 2 opening decisions — the earliest reconstructed gain and the earliest reconstructed loss, identified by when a position's mark moved between daily equity snapshots rather than by order sequence. All 4 of the selected decisions fall in Season 3, and all 4 are shorts.

The overlap is ADA. On Season 3's opening day, minutes apart, DeepSeek and GLM each opened a short on ADA, and by the next snapshot both of those ADA shorts had gained. The captured losses differ: DeepSeek shorted ADA and UNI together in a single cycle, and the UNI leg slipped to a loss on the next snapshot; GLM's reconstructed loss came from a separate short, on ARB, marked down a snapshot later.

These are 4 selected decisions, reconstructed from position-state changes between snapshots rather than trade fills — same-cycle round-trips never surface, and nothing in them explains the season-long gaps, since the pack logs no sizing, add-or-trim or hold-time. For completeness, the reported win rates: DeepSeek 22.9% to GLM's 20.0% in Season 3, 30.8% to 28.6% in Season 4, 75.0% to 39.1% in Season 5 — computed by season reports that count still-open positions among the trades, so no figure in that list is a closed-trade hit rate.

Recent pump to 0.267 rejected.

GLM-5GLM's first attributable gain: an ADA short it put on at the Season 3 open, marked higher a snapshot later.

RSI at 39.1 showing weakness.

GLM-5GLM's representative loss — a short on ARB it opened that same Season 3 morning; the next snapshot marked it down.

Season line-up: the model versions behind each result

SeasonDatesDeepSeek versionGLM versionAsset universeField
Season 3Mar–Apr 2026DeepSeek V3.2GLM-537 crypto assets9 models
Season 4Apr–May 2026DeepSeek V4 ProGLM-5.17 crypto assets9 models
Season 5May–Jun 2026DeepSeek V4 ProGLM-5.110 crypto assets10 models

How We Produced These Numbers

Every statistic here has a mechanical origin. Before this article ships, each figure must match the evidence pack, whose content hash pins every value to its source; the pack itself comes from a deterministic pass over each finished season's report, decision log and equity snapshots, and the paired results are frozen under that hash the moment a season is archived. A language model wrote these sentences; it supplied none of the values.

The trading itself is a simulation, worth stating outright. DeepSeek and GLM both see identical market data each cycle, form their own thesis, and submit their own orders into a paper account that fills at live prices; a 0.1% fee is charged per trade, while slippage, borrow and market-impact costs are not. Which conditions held fixed inside a season and which turned over between the 3 is what the limitations below take up; the narrower point here is that this ledger is rebuilt from live prices with fees modeled in, and was never settled at a real exchange.

Limitations and the Scoped Verdict

The load-bearing caveat is the sample. 3 shared completed seasons is 3 observations, and the pattern this piece is organized around — GLM's opening win followed by DeepSeek's 2 — rests entirely on them; another season could reorder it. Nothing here is a fixed trait of DeepSeek or GLM, because the model versions, prompts, tradable list and market outcomes all turned over between seasons; only the within-season conditions — the same capital, the 0.1% fee, the rulebook, the daily cadence — stayed matched.

Each metric carries its own edge. A return is marked to market with unrealized P&L folded in, so a headline can lean on open positions — DeepSeek's Season 5 figure did — which is what the realized/unrealized split above is for. The reported win rate counts positions that never closed, so it is not a closed-trade hit rate; read it beside that split. The opening decisions are rebuilt from daily-snapshot moves, not fills, so same-cycle round-trips stay hidden and season-end opens are marks rather than settlements. This was forward paper trading on live prices, not a backtest; slippage, market impact, borrow costs and real-capital risk were left out. Hold-time and profit factor stay out too — the archive holds no dependable value for either.

The verdict, scoped to the record: DeepSeek holds it, 2-1, having lost the opening shared season and won the 2 that followed, while GLM finished below zero in all 3. Whether any of that survives new model versions, new asset lists and new markets is precisely what 3 uneven seasons cannot say. Where DeepSeek and GLM stand in the competition still underway is on the live LLM trading benchmark; the DeepSeek vs GLM for trading evidence pack carries every number on this page, ready to be rechecked.

Frequently Asked Questions

In the DeepSeek vs GLM for trading matchup, which model came out ahead?

DeepSeek, on the season count. GLM won Season 3 — losing less in a season both models finished down — then DeepSeek won Season 4 and Season 5, for a 2-1 head-to-head across the shared seasons (Seasons 3–5). The paired gap, DeepSeek minus GLM, ran -3.99, +5.97 and +13.75 points, and GLM finished below zero in all 3. The running LLM benchmark tracks both models beyond these 3 seasons.

Did GLM ever finish a shared season in profit against DeepSeek?

No. Season by season, GLM closed at -7.67%, -0.57% and -1.90% in its 3 shared seasons with DeepSeek — under zero every time. Its lone head-to-head win, Season 3, came from losing less than DeepSeek (-7.67% to -11.66%), not from booking a gain, and its realized P&L was negative in all 3 seasons (-$721.59, -$65.46, -$488.16).

How did the DeepSeek and GLM return gap change across the 3 seasons?

It changed sign once, at the first transition. Measured as DeepSeek minus GLM, the gap read -3.99 points in Season 3 — GLM ahead, both models down — then +5.97 in Season 4 and +13.75 in Season 5. On season wins the 2 sides stood level after Season 4, and DeepSeek moved in front only with Season 5. Those 3 gaps are facts about 3 observations, not a trend to extend.

What does the GLM vs DeepSeek for trading record cover?

It runs across the 3 finished TradeRank crypto seasons both models completed together — Seasons 3–5 — each traded under one rulebook off an identical $10,000 stake. DeepSeek ran as V3.2, then V4 Pro; GLM as GLM-5, then GLM-5.1. GLM took Season 3; DeepSeek answered with Season 4 and Season 5 to settle the head-to-head 2-1.

Which DeepSeek and GLM versions traded in each season?

The builds changed once on each side. In Season 3, DeepSeek V3.2 faced GLM-5 across a 37-asset universe in a 9-model field. For Season 4 and Season 5 it was DeepSeek V4 Pro against GLM-5.1, over a 7-asset board in Season 4 and a 10-asset, 10-model field in Season 5. With both sides upgraded partway through the run, no single version owns the outcome — it is one head-to-head replayed across 3 seasons as the builds turned over on each side.

How much should a 3-season DeepSeek and GLM comparison be trusted?

As far as 3 observations reach, no further. Within the sample the sequence is plain: GLM won the opening shared season, DeepSeek won the 2 after it, and the paired gap was negative only in that opener. What the sample cannot tell you is whether the result survives new model versions, a different asset list or a different market — all of which changed between these seasons. Treat it as what these 3 seasons returned, not a durable ranking of DeepSeek over GLM.

Season 7 is live

Watch the AI models trade in real time

12 AI models trading live. Every decision logged and explained. Follow the competition on the TradeRank.ai arena.

See the live leaderboard →
← Back to The Signal