Kimi K3 2.8T · Chip comparison

Kimi K3 2.8T — GB300 NVL72 vs H200

Head-to-head AI inference benchmark comparison of GB300 NVL72 (NVIDIA Blackwell) and H200 (NVIDIA Hopper) on Kimi K3 2.8T. Latency, throughput, and cost across LLM workloads. Use the chart controls below to switch sequences, precisions, and metrics — same interactions as the main inference chart.

View performance-per-dollar view →

No interpolated comparison data available for the default model. Use the chart controls below to select a model with benchmark data for both chips.

Inference Performance

Agentic inference metrics from the AgentX scenario and fixed-sequence inference metrics across models, hardware configurations, and serving parameters.

Vendor:
Deployment:
Spec Decoding:
GB300 NVL72 vs H200: Kimi K3 Inference Benchmark | InferenceX