The kernel, live

Size a workload. No signup.

Pick a model, a chip, and a throughput target — get three-constraint sizing (compute · memory · bandwidth), cluster topology, and monthly + 3-year TCO, recomputed live. This is the same deterministic kernel behind every Instant Sizing Report and every platform sizing flow.

Workload

live · no signup
Your mintokTM
$0.487/ M-tok
See full plan →7-flow optimization
Modehosted-api baseline
Bound byBandwidth
Fleet8 × TPU 8i (raw 2)

Sizing breakdown

3-constraint
Compute
1 chips
Memory
1 chips
Bandwidth
2 chips
Chips required
2
raw requirement
Topology (rounded)
8
supported pod size
Tokens / chip / day
47.40M
effective ceiling
Power draw
11 kW
incl. server overhead

3-year TCO

cloud · lease · capex · hosted-API
ModeMonthly3-yr TCO$/M-tok
Cloud (on-demand)$35.0K$1.26M$13.52
Lease (committed HaaS)
CapEx (owned)
Hosted-API baseline$1.3K$45.5K$0.487

Hosted-API row is the "do nothing — pay per token" comparison. Output billed at this model's representative hosted rate; input estimated at 25% of output throughput (chat-typical). Input-heavy workloads (RAG, doc analysis) will land slightly higher than shown.

Calculated. Now optimise.
The full Mintok platform — continuous $/M-token optimisation across every workload, every chip, every contract.
Sign up →

Want the full report — hardware options ranked, build-vs-buy break-even, shareable write-up? Instant Sizing Report — $699.