Skip to content

mistral large 4 spent 14x more than deepseek on one haunted room

Comic-style witch pouring green potion into a bubbling cauldron by a gothic window, with a jack-o'-lantern and candles creative ai

deepseek v4 pro built a working three.js halloween room for 4 cents. mistral large 4 built the same room from the same prompt for 61 cents. that is 14x the bill and more than twice the time.

large 4 is mistral's new 1-trillion-parameter mixture-of-experts model with 49 billion parameters active per token. it is in public preview, and mistral promises open weights for the end of october. we tested it on a real build job against three open models of the same size – deepseek v4 pro, kimi k2.6 and mimo v2.5 pro – and all four rooms render.

four models, one prompt, one haunted room: a 32-second tour.

the models

modellabtotal paramsactive paramsweights
mistral large 4mistral1t49bpreview
deepseek v4 prodeepseek1.6t49bopen
kimi k2.6moonshot ai1t32bopen
mimo v2.5 proxiaomi1.02t42bopen, mit license

all four ran through openrouter.

the task

a halloween room with four objects:

  • a snow globe
  • a witch's cauldron
  • a newton's cradle of skulls
  • a crystal ball

the camera flies from one object to the next. same prompt for all four models.

the setup

the model plans first – room layout, camera path, shot timing and a file manifest with no file over ~350 lines – then writes the project one file per request, with everything it already wrote in context. our own harness on openrouter, 100k tokens max per reply, reasoning effort high for all four, a file that doesn't fit gets continued from its last line.

every room is a folder of plain es modules, three.js 0.170 from a cdn, every mesh and texture generated in code – no models, no images, no hdrs.

results

cost

#modelcost
1deepseek v4 pro$0.04
2mimo v2.5 pro$0.16
3kimi k2.6$0.38
4mistral large 4$0.61

time

#modeltime
1deepseek v4 pro30m 09s
2kimi k2.646m 43s
3mistral large 472m 36s
4mimo v2.5 pro72m 44s

output tokens

#modeltokens
1deepseek v4 pro83,539
2kimi k2.6132,647
3mistral large 4207,667
4mimo v2.5 pro234,014

lines of code

#modelloc
1mistral large 43,467
2mimo v2.5 pro2,239
3deepseek v4 pro2,046
4kimi k2.61,868

observations

mistral large 4 wrote the most code, 3,467 lines across 14 files, and built the most detailed objects: a glass snow globe on a walnut base, a cast-iron cauldron on three legs with jars and a spellbook, an arched window with cobwebs in both corners. it is also the priciest: $0.61 for the room, and 72 minutes.

deepseek v4 pro is the cheapest and the fastest: 4 cents and 30 minutes for the whole room, with 83,539 output tokens, the fewest of the four.

kimi k2.6 has the cleanest wide shot of the whole table: the four objects, the window and the pumpkin piles on the floor all fit in one frame.

mimo v2.5 pro has the most atmosphere and the longest thinking: 72 minutes and 234,014 output tokens. warm candlelight everywhere, a row of candles along the table and a black cat on the windowsill.

our take

mistral large 4 builds the best-looking objects: its snow globe, cauldron and skull cradle are the ones you would put on a thumbnail. kimi k2.6 builds the room you can read at a glance, with the whole table in one frame. mimo v2.5 pro goes for mood over detail and lets the candlelight do most of the work. deepseek v4 pro's room is plainer and darker, but it is complete, and it cost 4 cents. mistral's extra detail comes at 14x that price.

sources

ON AIR · RADIO.THEHYPE.NEWS ↗ ai news radio — 24/7