Kimi K3 2.8T · Performance per Dollar
Kimi K3 2.8T — GB300 NVL72 vs H200 Performance per Dollar
Cost per million tokens of GB300 NVL72 (NVIDIA Blackwell) versus H200 (NVIDIA Hopper) on Kimi K3 2.8T. Owning-hyperscaler TCO normalized by output tokens — performance per dollar across LLM workloads. Pick the more cost-efficient SKU at every target interactivity level. Use the chart controls below to switch sequences, precisions, and metrics — same interactions as the main inference chart.
Chip pricing (owning hyperscaler): GB300 NVL72 $2.31/chip/hr · H200 $1.22/chip/hr. Source: SemiAnalysis Market July 2026 Pricing Surveys & AI Cloud TCO Model.

No interpolated cost-per-token data available for the default model on this chip pair. Use the chart controls below to select a model and precision with benchmark data for both chips.
Inference Performance
Agentic inference metrics from the AgentX scenario and fixed-sequence inference metrics across models, hardware configurations, and serving parameters.
Vendor:
Deployment:
Spec Decoding: