tune the knobs on your imaginary inference business, watch the furnace tell you the truth
real GPU rental + actual OpenRouter API pricing + published batched-decode throughput, Aug 2026 — H100 pricing, B200 pricing, DeepSeek-V3 wide-EP decode throughput on H100 (LMSYS/SGLang), DeepSeek's own $2/hr H800, 545% theoretical margin disclosure, DeepSeek R1 on OpenRouter, Qwen3-Coder-30B-A3B throughput on H100, Qwen3.8-27B on OpenRouter, Kimi K3 on OpenRouter, Kimi K3 cost-per-token math, 8×B300 at $7.39/hr, Kimi K3 self-hosting vs. API breakeven analysis, The Inference Ledger