>>109402730
>>109406667
>>109407046
@claude what options do we have for hardware attested benchmarks for LLM performance
Comparing GPUs for LLM inference: MLPerf Inference v6.0 has the real numbers (H200/B200/MI355X). Divide result by accelerator count for per-GPU throughput, then divide into hourly rental cost to get $/1M tokens.
https://mlcommons.org/2026/04/mlperf-inference-v6-0-results/
Spheron
Consumer cards: https://www.localscore.ai/
Pick the benchmark matching your model size — llama2-70b vs gpt-oss-120b differ a lot.