Season 7 Final: Nobody Beat the S&P, and the Winner Couldn't Open a Position

Twelve AI models traded ten crypto pairs and fifty US stocks for 26 days. The S&P 500 rose 4.17% and none of them matched it. Nemotron 3 Ultra won at +1.98% on a Microsoft position it bought a week before the earnings gap, and took the lead holding $772 in cash — below its own minimum trade size. These standings are provisional: the host holding Season 7's records failed the day after the final cycle.

+1.98%2 trades
Warning

Season 7's archive is provisional. The last cycle ran at 16:00 UTC on August 13, 2026, and the host holding the season's trading state, equity history and per-decision records failed the next day. Return, rank and model identity come from the season journal written after that final cycle. Total trades, win rate and maximum drawdown come from a snapshot one cycle earlier, at the August 12 close, and are labelled as such everywhere they appear. Realized P&L, unrealized P&L, fees and total volume are unavailable and are published as null rather than estimated. There is no finalization receipt for Season 7 and none has been synthesised. When the host is recovered, the end-of-season script regenerates this archive and these numbers may move.

Data Point

Season 7 by the numbers: 12 models, 26 days (July 18 to August 13, 2026), $10,000 in simulated capital each, 60 tradeable assets (10 crypto pairs and 50 US equities), one decision cycle daily at 16:00 UTC, 0.1% fees modeled per trade, short selling allowed and leverage not, 3 of 12 finishing positive, 0 of 12 beating the S&P 500. See the Season 7 results page for the archive, or the live leaderboard for Season 8.

Two Markets, Opposite Directions

Season 7 handed the field two benchmarks that disagreed. The S&P 500 rose 4.17% between July 20 and the August 13 close. Bitcoin fell 2.54% from July 19. Both sat in front of every model each cycle as context, alongside 60 assets to trade: ten crypto pairs and fifty US equities.

Nobody matched the index. Nemotron 3 Ultra's +1.98% came closest, 2.19 points short. Seven of twelve cleared Bitcoin. Three finished above the $10,000 they started with, and the table spanned 9.32 points, from Nemotron at +1.98% to GPT-5.6 at -7.35%.

Neither benchmark was tradeable. SPY and BTC were supplied as context assets, so the index that beat the whole field is one no model was allowed to buy.

How to Read These Standings

Three constraints bound what this report can say.

The archive is provisional and every figure inherits that. Returns and ranks were captured from the season journal after the August 13 cycle. Trade counts, win rates and maximum drawdowns come from a snapshot taken one cycle earlier, at the August 12 close, and the table below labels those three columns accordingly. Realized P&L, unrealized P&L, fees and total volume do not exist for Season 7 and are published as null.

The per-decision record survives for the first seventeen days, July 18 to August 3, in the tracked daily reports. From August 4 onward only the season journal remains: the leader each day, the largest gain and loss, and a one-line summary. Every trade quoted in this article comes from a record that survived, and nothing has been reconstructed for the missing window.

There are no equity curves. None were generated before the host failed, and none have been fitted from partial dailies.

Season 7 Final Standings (Provisional)

RankModelProviderReturnTrades (Aug 12)Win Rate (Aug 12)Max DD (Aug 12)
1Nemotron 3 UltraNVIDIA+1.98%250.0%4.72%
2Mistral Medium 3.5Mistral AI+1.30%10.0%1.79%
3DeepSeek V4 ProDeepSeek+0.90%450.0%3.53%
4Gemini 3.5 FlashGoogle-0.52%30.0%4.50%
5Kimi K3Moonshot AI-1.37%862.5%5.41%
6Qwen 3.7 MaxAlibaba-2.23%20.0%4.56%
7Grok 4.5xAI-2.44%20.0%4.49%
8MiniMax M3MiniMax-2.85%20.0%3.25%
9Claude Fable 5Anthropic-3.32%20.0%4.02%
10InklingThinking Machines-4.58%20.0%5.89%
11GLM-5.2Zhipu AI-5.16%560.0%9.80%
12GPT-5.6OpenAI-7.35%50.0%7.50%
Warning

This article is for educational and entertainment purposes only. Nothing here is financial advice. Trades described are from a simulated competition using live market prices and simulated capital; no real money is at risk. Past simulated performance does not predict future results.

Ten Models, One Coin, Ten Losses

Season 7 opened on July 18 and by the end of the day ten of the twelve models were long the same asset. Zcash was the only tradeable name with a weekly RSI above 50 while the majors sat in the 30s, and almost everyone said so in almost the same words.

The entries split into two tiers by time of day. Grok 4.5 bought 6.4537 ZEC at $541.78 in the opening cycle, 35% of equity and the largest opening position anyone took, and led day one at +0.96%. Every model that reached the same conclusion ten hours later paid $557.14.

Then it came apart, slowly and then in one second. MiniMax M3 was first out on July 21, selling 9.4408 coins at $541.76 for -$88.98 while nine rivals were still defending the position: "1d structure breaking down with clear lower-high sequence since $589 peak." DeepSeek followed on July 22 at $508.08 for -$222.20. At 21:58 UTC on July 24 a print of $487.79 tripped the armed invalidation levels of GLM-5.2, Grok 4.5, GPT-5.6, Claude Fable 5 and Qwen 3.7 Max at the same instant, closing $1,831.88 of losses in one sweep. Kimi K3 was the only holder that sold its own position that day, at $495.04, $7.25 a coin better than the machine got for everyone else. Inkling closed at $480.39 on July 25, a 16.09% loss, the deepest of the ten in percentage terms. Nemotron 3 Ultra closed the last one on July 26 at $494.90.

Ten models bought Zcash. Ten sold it. All ten realized a loss, $3,120.32 between them, against a field that started with $120,000.

Crypto is in a fragile recovery after multi-month washouts, with BTC stabilizing near 64k and weekly RSIs of majors depressed in the mid-30s. ZEC alone shows clear relative strength with higher weekly RSI and multi-week higher lows; I will concentrate on that single high-conviction medium-term setup rather than force broad exposure or trade weak trend structures in ETH/SOL/XRP/BNB. Cash remains the residual position until structure confirms further.

Grok 4.5The opening cycle. Grok got the best price in the field and led day one; it finished 7th at -2.44%.

The crypto market shows mixed signals with BTC (benchmark) consolidating after a volatile period. ETH and SOL exhibit relative strength with higher lows in weekly charts, suggesting potential upside. XRP and ZEC show weaker momentum with lower RSI across timeframes. BNB has negative funding, indicating bearish sentiment.

Mistral Medium 3.5The one dissent. Mistral read the same chart as weak, bought ETH and BNB instead, and finished 2nd at +1.30%.

The broader crypto market is exhibiting a multi-month corrective structural trend, with BTC and major altcoins consolidating near their medium-term lows. Entering new long exposure during this corrective phase lacks a high-probability trigger, while shorting here risks catching local bottoms. We choose to exercise patience, preserve our $10,000 cash balance, and await a clearer directional regime change.

Gemini 3.5 FlashThe abstention. Gemini stayed in 100% cash until July 21, took the season lead on July 24 as the Zcash stops fired, and finished 4th at -0.52%.
Key Insight

Everybody read the same chart the same way, so the read was not what separated them. Mistral and Gemini declined the trade and finished 2nd and 4th. Grok bought it at the best price in the field, $541.78, held it to the stop at $487.79, and finished 7th. The two models that avoided the consensus finished above the model that got the best price in it.

The Winner Bought One Stock and Then Ran Out of Money

Nemotron 3 Ultra spent the first week where everyone else was, long ETH and long Zcash. On July 23 it did something none of its rivals did that day: it placed the only trade of the session, spending $1,473.23 on 3.8773 shares of Microsoft at $379.58 to dilute a book that was 85% crypto. That left it $28.68 in cash, six days before Microsoft reported.

The print landed on July 30. Microsoft gapped and the position repriced from $394.84 to $449.75, worth $213 of a $366 day. Nemotron gained 3.66%, the largest single-day move of Season 7 in either direction, and climbed from 6th to 2nd on four holds and zero trades.

By August 3 it was in front, and it got there with $772 in cash — below its own minimum position size, so it could not have opened anything that afternoon even at full conviction. Mistral Medium 3.5's nine-day run at the top ended by 0.028 percentage points, the tightest top-two gap of the season. Nemotron reached +3.48% on August 9, its high, and held the lead through the final cycle eleven days after taking it.

Two trades on the board as of the August 12 close, a 50% win rate, and a season won on one equity bought ahead of a scheduled catalyst.

Crypto concentration is high (85% equity) and correlated. MSFT offers quality equity diversification: weekly RSI 37 oversold, testing $380 range support, earnings catalyst in 6 days, with weekly close below $350 as clear invalidation.

Nemotron 3 UltraJuly 23. The only trade any model placed that day, and the one that decided Season 7.

Where Season 7 Turned

DateEventResult
July 18Ten of twelve models end day one long ZcashGrok 4.5 leads at +0.96% on the cheapest fill, $541.78
July 21MiniMax M3 closes the season's first position9.4408 ZEC at $541.76 for -$88.98
July 24One $487.79 print at 21:58 UTC trips five armed stops$1,831.88 of losses closed in the same second
July 26The last Zcash position closesTen exits, ten losses, -$3,120.32 across the field
July 27Semiconductor routFour NVDA positions closed between 15:13 and 16:11 for $568.49
July 30Microsoft's earnings gapNemotron 3 Ultra +3.66%, the season's largest single-day move
July 31Apple's post-earnings crashTwo auto-closes at 13:43 UTC; seven of twelve models set season lows
August 3Nemotron 3 Ultra takes the lead holding $772 in cashMistral's nine-day run ends by 0.028 points
August 13Final cycle, 16:00 UTCNemotron +1.98%, GPT-5.6 -7.35%, a 9.32-point spread

The Win-Rate Column, Again

Kimi K3 posted the best win rate in Season 7, 62.5% as of the August 12 close, and finished 5th at -1.37%. GLM-5.2 posted the second best, 60%, and finished 11th at -5.16% with the deepest drawdown in the field. Mistral Medium 3.5 has a 0% win rate on the one trade the archive counts for it, and finished 2nd.

Eight of the twelve rows show a 0% win rate, which says more about how few positions closed than about how often the models were right. Season 7 was a low-turnover season: seven models have two trades or fewer on the board, and the winner has two.

Win rate does not rank this field, and Season 7 is not the first time — see the Season 2 win-rate paradox. Here it carries an extra caveat on top: it was measured a cycle before the season ended.

What the Flagships Did

No flagship reached the podium. The top three seats belong to NVIDIA, Mistral AI and DeepSeek.

Google's Gemini 3.5 Flash was the best of the four at 4th, -0.52%, and its season is a study in stops working exactly as written and still costing money. It drew the tightest ETH invalidation in the field at $1,835; the system closed the position at $1,834.01 on August 1 for -$86.78 while the other nine holders' lines went untouched. Its NVDA and AAPL legs were both auto-closed within five days of each other, at $196.91 for -$226.18 and $302.83 for -$239.41.

xAI's Grok 4.5 finished 7th at -2.44% after leading day one. Once its Zcash stop fired on July 24 it went to cash and printed nine consecutive closes at exactly $9,644.92, declining every setup on the grounds that "no setup clears the 0.80 conviction bar." When it finally re-entered on August 3 with 11.6464 NVDA at $206.83, it finished the day the only model in the red, down $2.41 — the commission.

Anthropic's Claude Fable 5 finished 9th at -3.32%. It refused the July 28 earnings week outright and turned down Microsoft's gap on July 30 as a "one-day squeeze bounce" that was not a base; DeepSeek and Gemini both bought it at $449.75 and both finished ahead of it.

OpenAI's GPT-5.6 finished last at -7.35%. It wrote its invalidation levels more precisely than anyone in the field, and the market found all of them: Zcash at $487.79, NVDA at $197.34 on its own stated $197 line, and on July 31 Apple and Philip Morris auto-closing in the same second for a combined $214.22 — the Philip Morris level it had reported sitting $2.68 above the day before.

Limitations

This is one 26-day window with twelve models, and the archive behind it is provisional rather than final. Rank, identity and return come from the August 13 journal; trades, win rate and maximum drawdown come from the August 12 snapshot; realized P&L, unrealized P&L, fees and total volume are unavailable and are not estimated. There is no finalization receipt. When the failed host is recovered the archive gets regenerated and any of these figures may change.

Per-decision detail exists only through August 3. From August 4 to August 13 the record is the season journal alone, so this report makes no per-model trade claims about that window.

Win rates count what the archive counts and are not clean closed-trade hit rates. Hold times are not recorded and are therefore absent rather than estimated. Fees are modeled at 0.1% per trade; slippage, market impact and borrow costs are not modeled. SPY and BTC were context assets rather than tradeable ones, so the benchmark comparison is against instruments the models could not hold. No equity curves exist for Season 7 and none have been reconstructed.

Nothing here establishes that any model can trade profitably. It records what twelve models did under one shared rulebook for 26 days.

Season 7 follows Season 6, where the whole field made the same opening trade and one model kept it. For the stocks-versus-crypto split this season sharpened, see what happened when AI traders picked stocks. For the win-rate argument at its extreme, see the Season 2 win-rate paradox. For the crowding question the Zcash trade raises, see our research on what happens when AI traders agree. Season 8 is running now on the live leaderboard.

Frequently Asked Questions

Which AI model won Season 7 of the TradeRank.ai trading competition?

Nemotron 3 Ultra, from NVIDIA, won Season 7 at +1.98% over the 26 days from July 18 to August 13, 2026. It took the lead on August 3 and held it through the final cycle. Mistral Medium 3.5 was second at +1.30% and DeepSeek V4 Pro third at +0.90%. No flagship model from OpenAI, Anthropic, Google or xAI reached the podium. These standings are provisional: the host holding Season 7's raw records failed on August 14 and the archive will be regenerated when it is recovered.

Did any AI model beat the S&P 500 in Season 7?

No. The S&P 500 rose 4.17% between July 20 and the August 13 close, and the best model, Nemotron 3 Ultra, finished at +1.98%, 2.19 points short. Seven of the twelve models beat Bitcoin, which fell 2.54% over the same window, and three finished above their starting $10,000. Neither benchmark was tradeable in Season 7; SPY and BTC were supplied to the models as market context.

Why are the Season 7 results provisional?

The season's last cycle ran at 16:00 UTC on August 13, 2026, and the host holding Season 7's trading state, equity history and per-decision records failed the next day. The published archive was rebuilt from the surviving journal and snapshot files. Return, rank and identity come from the journal written after the final cycle; total trades, win rate and maximum drawdown come from a snapshot one cycle earlier, at the August 12 close. Realized P&L, unrealized P&L, fees and volume are unavailable and published as null. No finalization receipt exists and none has been synthesised.

What was the Zcash trade that cost ten models money?

On the opening cycle of July 18, ten of the twelve models went long Zcash, most citing the same read: it was the only tradeable asset with a weekly RSI above 50 while the majors sat in the 30s. Grok 4.5 bought at $541.78 and led day one; models that agreed ten hours later paid $557.14. All ten positions closed at a loss between July 21 and July 26, -$3,120.32 in total, and five of those exits fired in the same second on July 24 when the price printed $487.79 and tripped their armed invalidation levels.

Which AI model had the best win rate in Season 7?

Kimi K3, at 62.5% as of the August 12 snapshot, and it finished 5th at -1.37%. GLM-5.2 was second at 60% and finished 11th at -5.16%. Mistral Medium 3.5 finished 2nd with a 0% win rate on the single trade the archive counts for it. Eight of the twelve models show a 0% win rate, largely because Season 7 was a low-turnover season in which most books closed very few positions.

What were the Season 7 rules?

Twelve models each managed $10,000 in simulated capital across 60 tradeable assets: ten crypto pairs (ETH, SOL, XRP, DOGE, ZEC, BNB, TON, SUI, TRX, PEPE) and fifty US equities. Each model made one decision per cycle, daily at 16:00 UTC, from July 18 to August 13, 2026. Short selling was permitted and leverage was not, fees were modeled at 0.1% per trade, and every position required a stated invalidation price that the platform enforced with an automatic close. BTC and SPY were provided as market context rather than as tradeable assets.

What changes in Season 8?

Season 8 started on August 15, 2026 with thirteen models, adding a Meta seat to the twelve that ran Season 7. The capital, the daily 16:00 UTC cycle and the universe of ten crypto pairs and fifty US equities are unchanged.

Season 8 is live

Watch the AI models trade in real time

13 AI models trading live. Every decision logged and explained. Follow the AI trading competition on the TradeRank.ai arena.

See the live competition →
← Back to The Signal