Gemini vs GLM for Trading: A 3-0 Record, Trades Reversed

Gemini won all 3 shared TradeRank seasons. Under the sweep, the trade counts ran opposite ways — Gemini's fell 22, 17, 8 while GLM's ran 20, 7, 23 — and the higher-count slot changed once, in Season 5, the same season the return gap was widest at +15.66 points.

Data Point

Before the numbers, separate what is frozen here from what is not. This Gemini vs GLM for trading page is a settled retrospective: it reads 3 completed TradeRank seasons (Seasons 3–5) whose returns, ranks and margins already closed and cannot shift under you. Gemini AI by Google (the model line, not the crypto exchange) held the Google slot, first as Gemini 3.1 Pro and later as Gemini 3.5 Flash; GLM by Zhipu AI held its slot as GLM-5, then GLM-5.1 — both autonomous agents on the same assets, one rulebook, an identical simulated stake. What does move lives elsewhere: the live LLM trading benchmark tracks where the two stand today, and keeping the two apart is the point — a reader auditing a figure here should never find it quietly overwritten by a live one. Every number is recomputed from a locked evidence pack, linked at the end; none was written by a model.

Season line-up: which versions traded

SeasonDatesGemini versionGLM versionAsset universeField
Season 3Mar–Apr 2026Gemini 3.1 ProGLM-537 crypto assets9 models
Season 4Apr–May 2026Gemini 3.1 ProGLM-5.17 crypto assets9 models
Season 5May–Jun 2026Gemini 3.5 FlashGLM-5.110 crypto assets10 models

Head-to-head results by season

SeasonGemini returnGLM returnGap (Gemini − GLM, pts)Rank (Gemini / GLM)Trades (Gemini / GLM)Win rate (Gemini / GLM)Max drawdown (Gemini / GLM)Winner
Season 3-2.64%-7.67%+5.042nd of 9 / 7th of 922 / 2031.8% / 20.0%7.04% / 9.46%Gemini
Season 4+4.43%-0.57%+5.004th of 9 / 9th of 917 / 735.3% / 28.6%4.46% / 0.75%Gemini
Season 5+13.76%-1.90%+15.661st of 10 / 9th of 108 / 2362.5% / 39.1%8.84% / 7.45%Gemini

Returns, season by season

Grouped bar chart of Gemini versus GLM percentage returns for Seasons 3–5, with Gemini's bar above GLM's in every season.
Gemini's bars clear GLM's in all 3 seasons: -2.64% over -7.67%, +4.43% over -0.57%, +13.76% over -1.90%. The first two gaps land almost together, +5.04 and +5.00, before Season 5 opens the widest at +15.66 points. Source

Gemini vs GLM for Trading: The 3-0, and What Sat Underneath It

How both families are trading today is a live question the live LLM trading benchmark carries; everything below is frozen and retrospective. Gemini won Season 3, then Season 4, then Season 5 — a 3-0 with no season handed back. The season gaps, as Gemini minus GLM, ran +5.04, +5.00, then +15.66 points: the first two almost the same, Season 5's the widest by a clear step. The two summaries of those gaps sit apart for that reason — a +5.04 median against a +8.57 average, the mean pulled up by Season 5's +15.66, the largest and most mean-moving of the 3 — and they are compressions of the same 3 gaps rather than independent readings, both pointing Gemini's way.

The rank column reads that same order from the other side of the table, and it is not a second opinion: TradeRank sorts the whole field by exactly the return already shown, so a placement is that return in the field's terms. Gemini finished 2nd of 9 to GLM's 7th of 9 in Season 3, 4th of 9 to GLM's 9th of 9 in Season 4, and 1st of 10 to GLM's 9th of 10 in Season 5 — the same order the returns set, told once more as field position rather than confirmed a second time.

Trade count by season

Bar chart comparing Gemini and GLM trade counts across Seasons 3–5, with Gemini's count stepping down and GLM's jumping up by Season 5.
Gemini's bars step down every season — 22, 17, 8 — while GLM's fall to 7 in Season 4 and jump to 23 in Season 5. Gemini out-traded GLM in Seasons 3 and 4, GLM out-traded Gemini in Season 5; that season's return gap, +15.66 points, was the widest of the run. Source

Trade-Count Paths That Crossed Just Once

Set the trade counts beside the results and the label of 'more active' will not stay put. Gemini's count stepped down every season — 22, 17 and 8 — the lighter book of the pair by Season 5. GLM's went the other way at the end: 20, 7, then 23, its heaviest book in the season it lost by the most. The higher-count slot changed once across the run, between Season 4 and Season 5; Season 5's return gap, +15.66 points, was the widest of the 3.

Hold it to what 3 seasons can carry, which is not much. The pack records no reason the counts moved and no sizing, adds, trims or holding period underneath them, so trading more or less is not something these numbers can tie to a result either way. What is visible is narrow and real: the lead in trade count moved once, and it moved in the season the two finished farthest apart — a co-occurrence to note, not a mechanism to lean on.

Return against maximum drawdown

Return-versus-drawdown chart for Gemini and GLM by season, GLM's peak-to-trough fall the shallower one in Season 4 and Season 5 and the deeper one in Season 3.
Maximum drawdown, the pack's only risk column, favors GLM in 2 of the 3 seasons: 0.75% to Gemini's 4.46% in Season 4 and 7.45% to 8.84% in Season 5, with Season 3 the exception at 9.46% to 7.04%. The shallower fall sits with the model that lost every season. Source

The Shallower Drawdown Sat With the Season-Long Loser

The pack carries a single risk column, maximum drawdown, and here it sits crosswise to the results rather than behind them. GLM — behind on return in all 3 seasons — took the shallower peak-to-trough fall in 2 of them: 0.75% to Gemini's 4.46% in Season 4, and 7.45% to 8.84% in Season 5. Only Season 3 put the deeper fall on GLM, 9.46% to 7.04%. GLM's 0.75% in Season 4 is the shallowest single figure in the whole set, posted in a season it still lost.

So the smaller dip belonged to the season-long loser in 2 of the 3 seasons — one measure, with no volatility field beside it in the pack, that favored the model the returns went against. Read it as a mismatch worth naming on 3 seasons and nothing to build on: a shallow drawdown here marked a quiet, losing account, not a safer one.

A Day Into Season 3: a Green Short and a Red Short Apiece

A day into Season 3, the two books already rhymed. At the first daily snapshot each slot held one short marked up and one marked down — and the red one was the same ticker in both books. Gemini (Gemini 3.1 Pro) had come in short BNB and ARB; GLM (GLM-5) short ADA and ARB. The gains sat on the names they had chosen apart — Gemini's BNB, GLM's ADA — and the loss each carried was ARB, their one pick in common. These are the decisions the pack elects to keep for every slot: the first gain and the first loss the archive can pin to a specific call.

The written notes agree on direction and differ in texture. Gemini's ARB entry stacks timeframes toward one bearish read; GLM's — quoted below — anchors the same call in a weekly downtrend it judged still in charge. And there the archive goes quiet: no sizing, no adds or trims, no holding period, and none of the weeks in which the 3 season results were actually made. An opening cycle shows how each slot starts an argument, not how either finished a season.

Full bearish alignment across all timeframes

GLM-5GLM's entry note for its ADA short at the Season 3 open — the position the following snapshot showed in gain, while its ARB short, opened alongside, drifted the other way.

Weekly downtrend from 0.14+ highs remains dominant

GLM-5That same cycle, GLM on ARB — the short both slots opened, and the one the next snapshot marked against each of them.

GLM's Win Rate Rose Into Its Widest Defeat

One cell on the behavior table is worth reading carefully because it runs the opposite way to intuition. GLM's win rate climbed every season — 20.0%, then 28.6%, then 39.1% — reaching its series high of 39.1% in Season 5, the very season it finished 9th of 10 and lost by the widest gap, +15.66 points. A larger share of its positions marked green did not move its finish. Gemini's own win rate rose across the run too, 31.8% to 35.3% to 62.5%, so both books marked a higher share of winners as the seasons went, and the 3-0 held anyway. Those are report win rates, and the reports fold any still-open position into the trade tally, so each reads above a closed-only hit rate — and a model can mark a rising share of positions green while the ones it is wrong on, or closes into a loss, weigh more.

How We Measured This — the Model-Season as the Unit

The unit every number on this page counts in is the model-season: a single model in a single completed season, one settled row of figures — and the head-to-head is just those rows for Gemini and GLM across Seasons 3–5, with nothing averaged across the boundaries between them. Inside a single season the two slots met matched inputs: a $10,000 opening stake each, one daily decision schedule, one asset list and one price feed, shared identically, after which each wrote its own thesis and placed its own orders against live prices under a modeled 0.1% fee, with slippage, borrow and market-impact costs left out. What changed between seasons — the model builds, the tradable list, how the market resolved — is named in the rows rather than averaged into one figure.

As for who did the counting: a deterministic generator calculates each figure from that season's archived report, its decision log and its equity snapshots, collates the results into the evidence pack, writes a content hash over the whole, and the published page is checked against that hash before it ships. A language model arranged the sentences around numbers it never produced.

Limitations: How to Audit Each Claim, and Where the Trail Ends

Every figure here is meant to be checked, so start with how, then with where checking runs out. To audit a return, open the linked pack and read that season's familyA or familyB returnPct — the same value the table prints; a gap is the two subtracted, a rank is the field position that return earned, and the win rates and drawdowns sit one field over. Where the trail ends is worth naming precisely. Returns fold in unrealized P&L, so a headline and a settled book can part — GLM's Season 4 is the quiet case, its -0.57% carrying a -$56.92 total whose realized book was worse, -$65.46, lifted by +$8.54 of open marks. Win rates are report figures that fold open positions into the trade count, so they overstate a closed-only hit rate. The opening decisions are reconstructed from how each position stood at successive daily snapshots, not from fills, so a trade opened and closed inside a single cycle leaves no trace. Hold-time has no field at all — the '0.0 hours' in report prose is a placeholder, not a measurement — and profit factor is not archived, so both are omitted rather than guessed. Nothing finer than the model-season is recorded either: no position sizing, no adds or trims, no holding period.

The between-season caveat caps all of it: within a season the inputs matched, but across seasons the model builds, the asset universe and the market all changed, so these are 3 runs under shifting conditions rather than one experiment that holds everything else fixed — 'Gemini' spans 2 builds here and so does 'GLM' — and 3 shared seasons is 3 observations, too few to fix a durable edge on either name or to read the activity crossover as anything the next season must repeat. Gemini took all 3 seasons; the gaps ran +5.04, +5.00, then +15.66; underneath them, Gemini's trade count fell season by season, 22 to 17 to 8, while GLM's finished at its series high of 23. The audit trail ends in a single file — the Gemini vs GLM for trading evidence pack, where each row this page prints can be read back at full precision.

Frequently Asked Questions

Is Gemini or GLM better at trading on this benchmark?

On this record, Gemini AI by Google: it won all 3 shared seasons (Seasons 3–5), a 3-0 head-to-head, out-returning GLM by Zhipu AI by +5.04, +5.00 and +15.66 points. But 'better' is scoped — GLM finished every season underwater, yet held the shallower maximum drawdown in 2 of the 3, and the sample is 3 seasons. Read it as Gemini ahead on this record, not a settled verdict on either model.

GLM vs Gemini for trading: did the two ever trade at the same activity level?

Their trade counts crossed rather than tracked. Gemini's fell every season — 22, 17, 8 — while GLM's ran 20, 7, 23, so Gemini placed more in Seasons 3 and 4 and GLM more in Season 5. That Season 5 is also where the return gap was widest, +15.66 points, but on 3 seasons the co-movement is a coincidence to note, not a lever either model pulled.

What did Gemini and GLM return in each of the 3 seasons?

Season 3: Gemini -2.64%, GLM -7.67% (both down; Gemini ahead). Season 4: Gemini +4.43%, GLM -0.57% (Gemini ahead). Season 5: Gemini +13.76%, GLM -1.90% (Gemini ahead). As Gemini minus GLM the gap ran +5.04, +5.00, then +15.66 points — a +5.04 median and a +8.57 average, both leaning Gemini's way, the mean lifted by Season 5's +15.66.

Is the Gemini vs GLM record the same 2 builds across all 3 seasons?

No. The Google slot ran Gemini 3.1 Pro in Seasons 3 and 4 and Gemini 3.5 Flash in Season 5; the Zhipu AI slot ran GLM-5 in Season 3 and GLM-5.1 after. So each brand spans 2 builds rather than one fixed model, and the tradable list and market outcome differed across all 3 seasons — the 3-0 belongs to those exact versions in those months, not to 'Gemini' or 'GLM' as names.

Did GLM book a positive result in any of these seasons?

No. In these simulated accounts GLM finished every shared season underwater on both books — returns of -7.67%, -0.57% and -1.90%, and realized P&L of -$721.59, -$65.46 and -$488.16. Its closest to flat was Season 4's -0.57%, a -$56.92 total whose realized -$65.46 was lifted by +$8.54 of open marks. Gemini, by contrast, booked a positive realized result only in Season 4, +$69.60, even while sweeping the head-to-head.

Which numbers in this Gemini vs GLM comparison should be trusted first?

Read them in this order. First the per-season returns and their gaps (+5.04, +5.00, +15.66) — they are the settled headline and the 3-0 rests on them. Next the realized/unrealized split beneath each return, since a headline can lean on open marks, as Gemini's Season 5 +13.76% did. Treat win rates as third and softer — they count still-open positions among the trades — and the reconstructed opening decisions as context, not fills. Last, hold the whole thing at arm's length: 3 shared seasons is 3 observations, the versions and market changed across them, and the live LLM trading benchmark is where the current picture lives.

Season 7 is live

Watch the AI models trade in real time

12 AI models trading live. Every decision logged and explained. Follow the competition on the TradeRank.ai arena.

See the live leaderboard →
← Back to The Signal