The kernel, live
Size a workload. No signup.
Pick a model, a chip, and a throughput target — get three-constraint sizing (compute · memory · bandwidth), cluster topology, and monthly + 3-year TCO, recomputed live. This is the same deterministic kernel behind every Instant Sizing Report and every platform sizing flow.
Workload
live · no signupYour mintokTM
$0.487/ M-tok
See full plan →7-flow optimization
Modehosted-api baseline
Bound byBandwidth
Fleet8 × TPU 8i (raw 2)
Sizing breakdown
3-constraintCompute
1 chips
Memory
1 chips
Bandwidth
2 chips
Chips required
2
raw requirement
Topology (rounded)
8
supported pod size
Tokens / chip / day
47.40M
effective ceiling
Power draw
11 kW
incl. server overhead
3-year TCO
cloud · lease · capex · hosted-API| Mode | Monthly | 3-yr TCO | $/M-tok |
|---|---|---|---|
| Cloud (on-demand) | $35.0K | $1.26M | $13.52 |
| Lease (committed HaaS) | — | — | — |
| CapEx (owned) | — | — | — |
| Hosted-API baseline✓ | $1.3K | $45.5K | $0.487 |
Hosted-API row is the "do nothing — pay per token" comparison. Output billed at this model's representative hosted rate; input estimated at 25% of output throughput (chat-typical). Input-heavy workloads (RAG, doc analysis) will land slightly higher than shown.
Calculated. Now optimise.
The full Mintok platform — continuous $/M-token optimisation across every workload, every chip, every contract.
Want the full report — hardware options ranked, build-vs-buy break-even, shareable write-up? Instant Sizing Report — $699.