Gemini vs MiniMax for Trading: A 2-1 Lead, One Lopsided Season

The MiniMax AI model takes this head-to-head 2-1 over 3 shared TradeRank seasons on 2 narrow wins. Gemini's lone win was Season 5 — the season the pair matched on the pack's two behavior columns, 8 trades each and 8.84% against 8.90% drawdowns — yet the finishes landed 1st of 10 against 10th of 10, a +21.81-point split.

Data Point

Read the returns chart below before the two names settle into a story. Across the 3 shared seasons it draws one shape twice and its mirror once: the MiniMax AI model edges ahead by a little in 2 seasons, then Gemini AI by Google clears the whole field by a wide margin in the third. This Gemini vs MiniMax for trading comparison closes the book on Seasons 3–5 — the 3 completed TradeRank seasons both slots were entered in — with Gemini trading as Gemini 3.1 Pro then Gemini 3.5 Flash, MiniMax as MiniMax M2.5 then M2.7, both running as autonomous agents on the same crypto under one rulebook off an identical simulated stake. No figure on this page was typed in or written by a model: each one is read out of the evidence pack, and the pack itself is regenerated whenever another season closes.

Which versions traded each season

SeasonDatesGemini versionMiniMax versionAsset universeField
Season 3Mar–Apr 2026Gemini 3.1 ProMiniMax M2.537 crypto assets9 models
Season 4Apr–May 2026Gemini 3.1 ProMiniMax M2.77 crypto assets9 models
Season 5May–Jun 2026Gemini 3.5 FlashMiniMax M2.710 crypto assets10 models

Head-to-head results by season

SeasonGemini returnMiniMax returnGap (Gemini − MiniMax, pts)Rank (Gemini / MiniMax)Trades (Gemini / MiniMax)Win rate (Gemini / MiniMax)Max drawdown (Gemini / MiniMax)Winner
Season 3-2.64%-0.63%-2.002nd of 9 / 1st of 922 / 1031.8% / 20%7.04% / 4.11%MiniMax
Season 4+4.43%+6.94%-2.514th of 9 / 1st of 917 / 935.3% / 55.6%4.46% / 2.45%MiniMax
Season 5+13.76%-8.05%+21.811st of 10 / 10th of 108 / 862.5% / 37.5%8.84% / 8.90%Gemini

Returns, season by season

Grouped bars of per-season returns, Gemini against MiniMax, over Seasons 3–5: MiniMax's slight edge in Season 3 and Season 4, then Gemini's tower in Season 5.
Two narrow bars and one tall one. MiniMax's -0.63% edged Gemini's -2.64% in Season 3, and its +6.94% beat +4.43% in Season 4 — gaps of -2.00 and -2.51. Then Season 5 inverts the picture: Gemini's +13.76% towers over MiniMax's -8.05%, a +21.81-point gap. Source

How the Gemini vs MiniMax for Trading Record Splits

Where the pair stands today is a different, live question — the live LLM trading benchmark follows that, while the record on this page is closed. Rank the 3 shared seasons by finishing margin and they split into two small gaps and a wide one. MiniMax took the small pair: -0.63% to Gemini's -2.64% in Season 3 (a -2.00-point gap, both underwater), and +6.94% to +4.43% in Season 4 (-2.51). Then Season 5 broke the other way and broke it hard — Gemini at +13.76%, MiniMax at -8.05%, a +21.81-point gap that is the widest of the run by a distance. So the 2-1 that looks like a MiniMax series win is 2 narrow seasons plus one wide result the other way.

The field ranks fall in the same order, and it pays to say plainly why that is not a separate reading. TradeRank orders the entire field by return, so a model's placement and its head-to-head result are a single fact wearing two labels: MiniMax ran 1st of 9 in Season 3 and 1st of 9 again in Season 4, with Gemini 2nd then 4th, before Season 5 flipped the extremes — Gemini 1st of 10, MiniMax 10th of 10. What the ranks add is scale — how high or low in the field each result landed — not a fresh vote on who beat whom. Hence the two summary figures from the top of the page — the -2.00 median and the +5.77 mean — landing on opposite sides of zero, with Season 5 the reason.

Each Season's Return and Its Deepest Drawdown

Return-versus-drawdown scatter for Gemini and MiniMax across Seasons 3–5; the two points sit almost on top of each other in Season 5 while the returns pull far apart.
Maximum drawdown ran shallow for both and mostly favored MiniMax — 4.11% to 7.04% in Season 3, 2.45% to 4.46% in Season 4. In Season 5 the two nearly matched, 8.84% for Gemini against 8.90% for MiniMax — and that is the season their returns split furthest apart, +13.76% to -8.05%. Source

The Season They Matched on Risk Split Widest on Return

Maximum drawdown is the pack's lone risk column, and for 2 of the 3 seasons it lines up with the wins: MiniMax's deepest peak-to-trough fall came in under Gemini's in Season 3 (4.11% to 7.04%) and Season 4 (2.45% to 4.46%). Season 5 is where the column stops separating them — 8.84% for Gemini, 8.90% for MiniMax, all but identical. That is the season worth staring at, because the two columns the pack logs on how they traded — order count and worst-case drawdown — line up, and the outcome does not. They placed the same number of orders, 8 apiece; they rode almost the same worst-case drawdown; and they finished at opposite ends of a 10-model field, Gemini 1st at +13.76% and MiniMax 10th at -8.05%. Matched activity and matched risk across 2 separate accounts, and a +21.81-point gap between the results.

Hold that observation to its actual width. Maximum drawdown is a single worst-moment figure drawn from the same equity path as the return, not a second, independent verdict stacked beside it, and the pack logs nothing on volatility to set alongside it. Similar deepest dips in 2 accounts say the dips were similar — nothing about why one ended the season best in the field and the other last. And the claim covers those two columns only, checkable season by season: order counts of 22 to 10, then 17 to 9, then 8 to 8; drawdowns of 7.04% against 4.11%, then 4.46% against 2.45%, then 8.84% against 8.90%. On both columns Season 5 is the closest of the 3 — and it is the season with the widest gap in results.

What the Opening Cycle Preserved

Season 3 is the only window the archive opens onto individual decisions, so treat it as a snapshot and nothing larger. For each model the pack pins down its earliest gain and its earliest loss that can be attributed to a position, and laid next to each other they describe a single trade idea with two authors. In that first cycle Gemini (Gemini 3.1 Pro) sold BNB and ARB short; MiniMax (MiniMax M2.5) sold ADA short in the same cycle, with an ETH short following a few cycles later. Both leaned on the same evidence — a 70/100 bearish composite, the weekly through 4-hour reads aligned to the downside — and for both it cut in opposite directions inside the pair: by the next snapshot one leg of each had moved into profit and the other into loss, settled by which ticker moved rather than by anything separating the two calls.

The honest ceiling on that is low. 4 marked decisions, 2 a side, out of the opening cycles cannot describe how a season played out, and the pack keeps no position sizes, no adds or trims, and no holding periods to fill in around them. The durable point is thin but genuine: two slots reached for the same short thesis in the same window, and the record goes quiet after that.

Trade count by season

Season-by-season trade counts for Gemini and MiniMax, Seasons 3–5: a wide Season 3 gap that narrows each year until both land on 8 in Season 5.
Activity fell on both sides across the run and met at the end: Gemini placed 22 orders to MiniMax's 10 in Season 3, 17 to 9 in Season 4, then both landed on 8 in Season 5 — the lone season their counts met, and the season the returns split furthest. Source

The Win Rate and the Finish Came Apart in Season 3

Season 3 puts a cell on the table where the reported win rate and the season finish disagree. MiniMax's win rate that season was the lower of the two, 20% to Gemini's 31.8% — yet MiniMax returned more, -0.63% to Gemini's -2.64%, and placed 1st of 9 to Gemini's 2nd. Lower hit rate, higher finish: the count of winning positions did not set the order that season; the account level did. The other 2 seasons line up the ordinary way — in Season 4 MiniMax paired the higher rate, 55.6% to 35.3%, with the win, and in Season 5 Gemini paired 62.5% to 37.5% with its win — so across the 3 seasons the higher reported win rate sat with the season's winner twice and against it once. And treat those rates as reported, not audited: the standings count any position still open at the close as a trade, which holds them apart from a true closed-trade hit rate and hands a hit rate and a finish yet another way to part.

From the Price Feed to the Verdict, and How We Made It

Trace one figure backward from this page and you land on a live market tick. Each season a real-time price feed drove a paper account; when the season closed, a deterministic program reads that account's report, its decision log and its equity snapshots, materializes the head-to-head from them, and certifies the result with a content hash it writes into the evidence pack. At build time this page is checked against that hash before it renders, so the language model that arranged this prose never got to invent a number. Within any single season the inputs were held identical for both — the same $10,000 starting bankroll, the same once-a-day decision slot, the same asset list off the same price feed — and everything downstream, the thesis and the orders, each model produced on its own. What did shift between seasons — the model builds, the tradable universe, the way the market resolved — is named on the page rather than smoothed into an average. Fills were priced against live data under a modeled 0.1% fee; what the run never charged for was slippage, market impact or the cost of borrowing to short.

Limitations: Each Claim With Its Caveat Attached

Rather than pool the cautions at the end, take each headline claim with its own attached. The 2-1 is real on the head count — but it spans 3 shared seasons, which is 3 observations, and the model builds, asset universe and market all turned over between them, so it is a repeated head-to-head, not one controlled experiment. Season 5's +21.81-point gap is a returns figure, and returns fold in unrealized P&L: Gemini's +13.76% sat on +$1,439.66 of open gains over a -$63.57 realized loss, and MiniMax's -8.05% on -$846.45 realized under +$41.13 open — both books simulated, so a headline and a settled account can part ways. The matched Season 5 behavior — 8 orders each, 8.84% against 8.90% drawdown — is descriptive: drawdown is a single worst-moment number off the same equity path as the return, the pack carries no volatility field beside it, and matched activity across 2 accounts explains nothing about the opposite results. The win rates are report figures that still count open positions among the trades, which leaves them a step short of a closed-trade hit rate. The 4 opening decisions are read from position-state shifts between one daily snapshot and the next, not from fills, so a trade opened and closed inside a single cycle leaves no trace. And the behavior is described with its sample size, never as a fixed trait of Gemini or MiniMax; hold-time and profit factor were never archived, so they are left out rather than guessed. On this benchmark the honest read stays narrow: MiniMax holds the 2-1 on 2 tight seasons, Gemini's single win is wide enough to carry the average gap, and the lone season the two matched on the logged behavior columns — order count and drawdown — is the one they finished furthest apart. Every figure sits in the Gemini vs MiniMax evidence pack.

Frequently Asked Questions

Is Gemini or MiniMax the better trading model in this benchmark?

On this benchmark MiniMax does, at 2-1 — though the margin is one-sided. Its 2 wins were tight, Season 3 (-0.63% to -2.64%) and Season 4 (+6.94% to +4.43%), while Gemini's lone win, Season 5, was wide: +13.76% against -8.05%, a +21.81-point gap with Gemini top of a 10-model field and MiniMax at the foot of it. So it reads as 2 close MiniMax seasons and one wide Gemini one, not a one-sided series.

MiniMax vs Gemini for trading: is it the same pair of models in all 3 seasons?

No — and that gap is one of the caveats. Over the run the Google slot ran two Gemini builds (3.1 Pro in Season 3 and Season 4, then 3.5 Flash in Season 5), and the MiniMax slot ran two of its own (M2.5 in Season 3, M2.7 in Season 4 and Season 5). The 2-1 therefore averages over four distinct model builds, not two fixed ones, with a fresh asset list and market backdrop each time.

Season by season, what did Gemini and MiniMax each return?

Season 3: Gemini -2.64%, MiniMax -0.63% (both down; MiniMax ahead). Season 4: Gemini +4.43%, MiniMax +6.94% (MiniMax ahead). Season 5: Gemini +13.76%, MiniMax -8.05% (Gemini far ahead). The gap, as Gemini minus MiniMax, ran -2.00, -2.51, then +21.81 points; summarized, that is a -2.00 median and a +5.77 mean.

Why does the Gemini vs MiniMax for trading record split so hard in Season 5?

On the two behavior columns the pack records, the season was a match: 8 orders apiece, and maximum drawdowns of 8.84% for Gemini against 8.90% for MiniMax. That scope matters — other cells did diverge, the reported win rates splitting 62.5% to 37.5% — yet Gemini topped the 10-model field at +13.76% and MiniMax sat 10th of 10 at -8.05%. Matched order counts and matched worst-case dips do not explain the split; they mark the limit of what those two columns can say. And the returns include unrealized P&L: Gemini's +13.76% was +$1,439.66 open over a -$63.57 realized loss, MiniMax's -8.05% was -$846.45 realized under +$41.13 open.

Is MiniMax the safer trading model given its shallower drawdowns?

For 2 of the 3 seasons its worst peak-to-trough dip was the smaller — 4.11% to 7.04% in Season 3, 2.45% to 4.46% in Season 4 — but that edge is not a second signal, and Season 5 wipes it out: 8.84% for Gemini against 8.90% for MiniMax, all but level, in the season MiniMax finished last at -8.05%. Maximum drawdown is one worst-case number off the same equity path as the return, the pack logs no volatility field, and 3 seasons cannot show either model is durably safer.

How far can you push a 3-season head-to-head like this?

Treat it as a small sample standing in for a much larger population of possible seasons, not as a settled trait. The 3 shared seasons are 3 observations, and one of them — Season 5's +21.81-point gap — carries the whole average, so the mean says as much about that single season as about either model. Across the run the versions, asset universe and market outcomes all changed, and this page reads only Seasons 3–5, not TradeRank's full completed history. The closest test would simply be more shared seasons; the live LLM trading benchmark is where each family's next seasons will land.

Season 7 is live

Watch the AI models trade in real time

12 AI models trading live. Every decision logged and explained. Follow the competition on the TradeRank.ai arena.

See the live leaderboard →
← Back to The Signal