mintok.ai · self-service
Instant Sizing Report
Describe your AI workload in plain English. Get hardware options, TCO, $/M-token, and break-even volume.
Prices have moved since this report: 9 price changes across 7 models since 2026-07-28.
This page is a snapshot. On the platform, these numbers reprice daily against the live catalogue.
Start a 14-day trial →Chatbot
Customer support chatbot . 100 users . inpute/output =100/200
Self-host options
| Model | Chip | Chips | Racks | Binding | CapEx | TCO/yr | $/M-tok | Response |
|---|---|---|---|---|---|---|---|---|
| Qwen2.5 3BT1 | MI500 | 1 | 1 | Bandwidth | $0.1M | $0.03M | $0.07 | 138 ms |
| Gemma 2 2BT1 | MI500 | 1 | 1 | Bandwidth | $0.1M | $0.03M | $0.05 | 98 ms |
| Llama 3.2 1BT1 | MI500 | 1 | 1 | Bandwidth | $0.1M | $0.03M | $0.02 | 59 ms |
Best $/M-token highlighted. Binding = which physical constraint sets the chip count.
Hosted-API comparison (cheapest per tier band)
| Band | Model | $/event | $/month at your volume |
|---|---|---|---|
| Mid (T5–6) | Kimi K3Moonshot | $0.00330 | $10 |
Build vs buy
Best self-host option: $2,500/mo · hosted on Kimi K3: $10/mo. Self-hosting breaks even at ≈24,907 events/day (you stated 100).
Assumptions
- · pue: 1.3
- · mbuPct: 50
- · mfuPct: 45
- · acqMode: CapEx
- · kwhCost: 0.07
- · utilPct: 70
- · batchSize: 8
- · amortYears: 4
- · peakFactor: 3
Go deeper
Validate model quality on your real prompts with a benchmark run, or design the agent layer with the Agent Cost Designer. Methodology: how we size.
Delivered by Mintok — AI infrastructure economics. This link is private to whoever holds it.