makemode
the models · eu-hosted

the models you build with.

MakeMode builds on a small set of open AI models — from the quickest to the most polished. Pick by what you need; every one runs on European infrastructure, so your work never leaves the EU whichever you choose. Here's how they differ, and how they measure up — including speed numbers we measure ourselves.

pick by what you need

fastest to best-looking.

The same trade-off you'll see in the model picker in the web app and demo: quicker models draft simpler pages; larger ones take longer but build more detailed, polished work. One choice, no wrong answer.

Fastest

Mistral Small 3.2

The quickest way to a first draft — great for simple pages and fast iteration.

Mistral · France
technical detail →

The technical specs for this model — size, context window, licence and how it's served in the EU — will appear here.

Balanced

Devstral 2

Mistral's agentic coding model — more detail, still quick. Strong on code-heavy builds.

Devstral · EU
technical detail →

The technical specs for this model — size, context window, licence and how it's served in the EU — will appear here.

Capable

Mistral Medium 3.5

The European flagship — detailed, well-built pages. A longer wait for higher-quality output.

Mistral · France
technical detail →

The technical specs for this model — size, context window, licence and how it's served in the EU — will appear here.

Best-looking

GLM 5.2

The most polished output — a strong open-weights reasoning model that writes the most detailed pages.

open weights · EU-hosted
technical detail →

The technical specs for this model — size, context window, licence and how it's served in the EU — will appear here.

Largest

Qwen3.5 397B

The biggest model here — deeply detailed builds, and the longest wait. For work that rewards patience.

open weights · EU-hosted
technical detail →

The technical specs for this model — size, context window, licence and how it's served in the EU — will appear here.

how to choose in opencode → — in the web app it's a picker; in the terminal it's one line of config. Same models, one key.

how they measure up

the benchmarks, honestly.

Benchmarks say what a model is good at — not where your data goes (that's identical for all of them). Here's a snapshot on coding; rankings move weekly, so we link the live boards.

coding capabilitySWE-bench Verified · reported July 2026
ModelBest forSWE-bench Verified
Mistral Small 3.2fastest · lighter draftsquick pages, simple toolsgeneral model
Devstral 2agentic coding, EUcoding-heavy builds72%
Mistral Medium 3.5EU flagshipdetailed, polished pages78%
GLM 5.2open weights, EU-hostedmost polished outputtop open-weights †
Qwen3.5 397Blargest · open weightsdeeply detailed buildstop open-weights tier †
On WebDev Arena (building web front-ends), Claude leads overall; among the open, EU-hostable models, GLM and the Qwen flagship tier rank highest.
† GLM 5.2 leads open-weight models on LMArena; its makers report 62% on the harder SWE-bench Pro (a tougher test, not comparable to Verified). See the live boards: LMArena · Artificial Analysis · SWE-bench.
measured on our rails

how fast, measured by us.

Leaderboards test models in a lab. This is what they do here — on the same European servers that build your pages. We run this benchmark ourselves, publish the raw data, and re-measure when the serving changes.

a full page, start to finish — same prompt to every modelmeasured 2 july 2026 · scaleway, paris · median of 3 streamed runs
Mistral Small 3.2starts writing in ~0.3s · 144 tok/s
13sfull build
Mistral Medium 3.5starts writing in ~0.3s · 67 tok/s
26sfull build
GLM 5.2thinks ~3s, then 212 tok/s — wrote 3× the page of the others
29sfull build
Devstral 2starts writing in ~0.3s · 67 tok/s
32sfull build
Qwen3.5 397Bthinks ~11s first, then 150 tok/s
43sfull build
How we measure: the same page-build prompt to every model, streamed, three runs, medians, no token cap — every build finishes naturally. "Full build" is wall time from request to last token; tok/s uses the provider's reported completion tokens (not chunk counting); "starts in" includes any thinking time, because that's the wait you actually see. Models write different amounts for the same ask — GLM wrote a ~5,700-token page where Mistral Medium wrote ~1,700 — so time and detail trade off. Speeds drift with provider load and serving changes — we measured the same model swing 2× between days — so we re-measure rather than assume, and route each model to whichever EU provider serves it fastest. Raw data: model-speed.json.
the footprint

greener by where it runs.

Running in Europe isn't only about your data — it's about the grid underneath. MakeMode's models run on European data centres chosen for lower-carbon energy. We're measuring the energy and CO₂ of a real build the same way we measure speed — openly, with the raw numbers published here.

energy & co₂ per buildmeasured on our rails · published openly
Per-model energy and CO₂ figures are being measured now and will appear here — like the speed benchmark above, with the raw data linked. See also the full environment story.
the cost

what a build costs.

Open models are far cheaper to run than closed frontier services — and we'd rather show you the price than hide it. Here's the cost per model, so you can pick by budget as well as by polish, and see exactly what you're paying for.

cost per modelprice per 1M tokens · eu-hosted
Per-model prices are being finalised and will appear here, alongside a typical cost per build. See also inference economics and unit economics.
the same for every model

whichever you pick, it stays in the EU.

Sovereignty isn't a property of the model — it's how MakeMode runs all of them: on European clouds under GDPR (Scaleway in France, with some models served from EU data centres in Finland), your words in, a page out, nothing leaving the EU. An open model is a file we run, not a service phoning home — so where its makers are doesn't change where your data goes. That's true for the fastest Mistral and the most polished GLM alike.

the full sovereignty story →

build with one now.

Describe a page, watch it build, publish a link — on European rails.

early-stage; data-residency and processing commitments described here are design intent and operating goal, becoming contractual as data-processing agreements are signed — not, at this stage, asserted as guarantees.