Measured workload economics
Verified GPU cost benchmarks
Convert exact MLPerf 6.0 results into published snapshot economics or enter your own rate. See cost per million reported tokens, cost per batch unit and cost to reach a published training quality target—without transferring the result to a different model or hiding the assumptions.
- Official MLCommons source paths pinned to full commits
- 60 raw training runs retained by SHA-256 hash
- 3 independent submitters with provider-specific official price matches
Benchmark price snapshot:Checking the declared provider-SLA boundary.Valid through ; historical after that instant.
Transparent repricer
Change the rate, never the published result
The benchmark throughput or time stays locked. Change only the per-GPU price and optional target volume to audit the arithmetic.
Exact published system
Published reference economics
These 12 results cover 3 published systems from 3 matched cloud providers. Each result was re-priced at price-catalog generation only with the same provider, exact normalized GPU model and on demand market, then multiplied by that benchmark's published GPU count. This is a transparent reference calculation, not proof of identical topology, capacity or performance for another workload. The contributing price snapshot is valid through and remains visible as historical evidence afterward.
qwen3-vl-235b-a22b
$0.2229 / 1K samples
- Published result
- 78.28 Samples/s
- Reference time
- 12.78 s / 1K samples
- System
- 8× NVIDIA B300-SXM-270GB
- Cloud price match
- Nebius AI Cloud · B300
- Reference cluster rate
- $62.80 / h
- Reference yield @ $100 compute
- 448.72K samples
- Purchase model
- on demand
- Price observed
- 2026-09-09 22:35 UTC
- Capacity
- 8 GPUs available together: not verified
- Evidence
- 1 verified result
- Source revision
4d3916ac9cf4
F1_HIERARCHICAL: 0.7880721729948121
Inspect source evidence ↗qwen3-vl-235b-a22b
$0.3864 / 1K queries
- Published result
- 45.15 Queries/s
- Reference time
- 22.15 s / 1K queries
- System
- 8× NVIDIA B300-SXM-270GB
- Cloud price match
- Nebius AI Cloud · B300
- Reference cluster rate
- $62.80 / h
- Reference yield @ $100 compute
- 258.83K queries
- Purchase model
- on demand
- Price observed
- 2026-09-09 22:35 UTC
- Capacity
- 8 GPUs available together: not verified
- Evidence
- 1 verified result
- Source revision
4d3916ac9cf4
F1_HIERARCHICAL: 0.7865237675409595
Inspect source evidence ↗deepseek-r1
$0.2517 / 1M reported tokens
- Published result
- 69,318.9 Tokens/s
- Reference time
- 14.43 s / 1M reported tokens
- System
- 8× NVIDIA B300-SXM-270GB
- Cloud price match
- Nebius AI Cloud · B300
- Reference cluster rate
- $62.80 / h
- Reference yield @ $100 compute
- 397.37M reported tokens
- Purchase model
- on demand
- Price observed
- 2026-09-09 22:35 UTC
- Capacity
- 8 GPUs available together: not verified
- Evidence
- 1 verified result
- Source revision
4d3916ac9cf4
exact_match: 81.58614402917047 TOKENS_PER_SAMPLE: 3722.1116681859617
Inspect source evidence ↗deepseek-r1
$0.2888 / 1M reported tokens
- Published result
- 60,413.4 Tokens/s
- Reference time
- 16.55 s / 1M reported tokens
- System
- 8× NVIDIA B300-SXM-270GB
- Cloud price match
- Nebius AI Cloud · B300
- Reference cluster rate
- $62.80 / h
- Reference yield @ $100 compute
- 346.32M reported tokens
- Purchase model
- on demand
- Price observed
- 2026-09-09 22:35 UTC
- Capacity
- 8 GPUs available together: not verified
- Evidence
- 1 verified result
- Source revision
4d3916ac9cf4
exact_match: 81.58614402917047 TOKENS_PER_SAMPLE: 3721.7894257064722
Inspect source evidence ↗gpt-oss-120b
$0.1632 / 1M reported tokens
- Published result
- 106,885 Tokens/s
- Reference time
- 9.36 s / 1M reported tokens
- System
- 8× NVIDIA B300-SXM-270GB
- Cloud price match
- Nebius AI Cloud · B300
- Reference cluster rate
- $62.80 / h
- Reference yield @ $100 compute
- 612.72M reported tokens
- Purchase model
- on demand
- Price observed
- 2026-09-09 22:35 UTC
- Capacity
- 8 GPUs available together: not verified
- Evidence
- 1 verified result
- Source revision
4d3916ac9cf4
exact_match: 82.959
Inspect source evidence ↗gpt-oss-120b
$0.1737 / 1M reported tokens
- Published result
- 100,437 Tokens/s
- Reference time
- 9.96 s / 1M reported tokens
- System
- 8× NVIDIA B300-SXM-270GB
- Cloud price match
- Nebius AI Cloud · B300
- Reference cluster rate
- $62.80 / h
- Reference yield @ $100 compute
- 575.75M reported tokens
- Purchase model
- on demand
- Price observed
- 2026-09-09 22:35 UTC
- Capacity
- 8 GPUs available together: not verified
- Evidence
- 1 verified result
- Source revision
4d3916ac9cf4
exact_match: 83.337
Inspect source evidence ↗GPT-OSS 20B pretraining
$247.5209 / reference run
- Published result
- 1,618.96 seconds
- Reference time
- 26.98 min to published quality target
- System
- 64× NVIDIA Blackwell GPU (B200-SXM-180GB)
- Cloud price match
- CoreWeave · B200
- Reference cluster rate
- $550.40 / h
- Reference yield @ $100 compute
- 0.4 reference runs
- Purchase model
- on demand
- Price observed
- 2026-09-09 22:35 UTC
- Capacity
- 64 GPUs available together: not verified
- Evidence
- 10 verified runs
- Source revision
eabf23a07b2a
C4 · 3.34 log perplexity
Inspect source evidence ↗
Llama 3.1 8B pretraining
$151.6928 / reference run
- Published result
- 992.18 seconds
- Reference time
- 16.54 min to published quality target
- System
- 64× NVIDIA Blackwell GPU (B200-SXM-180GB)
- Cloud price match
- CoreWeave · B200
- Reference cluster rate
- $550.40 / h
- Reference yield @ $100 compute
- 0.66 reference runs
- Purchase model
- on demand
- Price observed
- 2026-09-09 22:35 UTC
- Capacity
- 64 GPUs available together: not verified
- Evidence
- 10 verified runs
- Source revision
eabf23a07b2a
C4 · 3.3 log perplexity
Inspect source evidence ↗
GPT-OSS 20B pretraining
$87.3772 / reference run
- Published result
- 5,008.89 seconds
- Reference time
- 83.48 min to published quality target
- System
- 8× NVIDIA Blackwell Ultra GPU (B300-SXM-270GB)
- Cloud price match
- Nebius AI Cloud · B300
- Reference cluster rate
- $62.80 / h
- Reference yield @ $100 compute
- 1.14 reference runs
- Purchase model
- on demand
- Price observed
- 2026-09-09 22:35 UTC
- Capacity
- 8 GPUs available together: not verified
- Evidence
- 10 verified runs
- Source revision
eabf23a07b2a
C4 · 3.34 log perplexity
Inspect source evidence ↗GPT-OSS 20B pretraining
$86.0397 / reference run
- Published result
- 5,787.43 seconds
- Reference time
- 96.46 min to published quality target
- System
- 8× NVIDIA Blackwell GPU (B200-SXM-180GB)
- Cloud price match
- Lambda · B200
- Reference cluster rate
- $53.52 / h
- Reference yield @ $100 compute
- 1.16 reference runs
- Purchase model
- on demand
- Price observed
- 2026-09-09 22:35 UTC
- Capacity
- 8 GPUs available together: not verified
- Evidence
- 10 verified runs
- Source revision
eabf23a07b2a
C4 · 3.34 log perplexity
Inspect source evidence ↗
Llama 3.1 8B pretraining
$76.0547 / reference run
- Published result
- 5,115.79 seconds
- Reference time
- 85.26 min to published quality target
- System
- 8× NVIDIA Blackwell GPU (B200-SXM-180GB)
- Cloud price match
- Lambda · B200
- Reference cluster rate
- $53.52 / h
- Reference yield @ $100 compute
- 1.31 reference runs
- Purchase model
- on demand
- Price observed
- 2026-09-09 22:35 UTC
- Capacity
- 8 GPUs available together: not verified
- Evidence
- 10 verified runs
- Source revision
eabf23a07b2a
C4 · 3.3 log perplexity
Inspect source evidence ↗
Llama 3.1 8B pretraining
$75.3729 / reference run
- Published result
- 4,320.74 seconds
- Reference time
- 72.01 min to published quality target
- System
- 8× NVIDIA Blackwell Ultra GPU (B300-SXM-270GB)
- Cloud price match
- Nebius AI Cloud · B300
- Reference cluster rate
- $62.80 / h
- Reference yield @ $100 compute
- 1.33 reference runs
- Purchase model
- on demand
- Price observed
- 2026-09-09 22:35 UTC
- Capacity
- 8 GPUs available together: not verified
- Evidence
- 10 verified runs
- Source revision
eabf23a07b2a
C4 · 3.3 log perplexity
Inspect source evidence ↗Evidence boundary
What this proves—and what it does not
These are results for 3 exact published systems from3 independent MLPerf submitters, each with a named model, scenario, software stack, topology and quality result. The economics layer selects only the same provider and exact normalized GPU model, then multiplies that provider's per-GPU rate valid at price-catalog generation rate by the GPU count of the published system. After the displayed price-snapshot deadline, that derived result is historical evidence.
Machine-readable evidence