WORKLOAD TRADE-OFF MATRIX

GPU Comparison Engine

Select any two GPUs for a side-by-side evaluation of VRAM capacity, memory bandwidth, TFLOPS throughput, and current spot pricing across workload profiles.

Methodology: How GPU Comparison Works

OpenGPU Radar compares GPU instances across four dimensions: Price, Performance, Memory, and Availability. Pricing data is aggregated from Spheron, RunPod, Vast.ai, and Lambda Labs APIs, refreshed every 6 hours.

Cost per 1M Tokens = (Hourly GPU Cost) / (Tokens/Second ร— 1,000,000) Lower values indicate better economics. Throughput Index = Effective Tokens/Second (vLLM FP8, Llama 70B) Measured end-to-end output throughput including KV-cache retrieval. Example: H100 SXM5 at $1.89/hr delivers ~120 tok/s โ†’ $1.89 / (120 ร— 1M) = $0.000158 per 1M tokens

Price Data Source
Cloud Provider APIs
Refresh Rate
Every 6 hours
Baseline GPU
RTX 4090 ($0.34/hr)
Benchmarks
vLLM FP8 inference

Hardware Requirements by Workload

Minimum VRAM and GPU specs required for each workload category. Determined via entity-graph VRAM calculations.

WorkloadMin GPUMin VRAMPrecisionRecommended GPU
70B LLM InferenceNVIDIA H100 SXM580 GBFP8NVIDIA H200 SXM5 or NVIDIA B200 Blackwell
30B LLM InferenceNVIDIA L40S48 GBFP8NVIDIA A100 80GB SXM4
7B LLM InferenceNVIDIA GeForce RTX 409024 GBFP16NVIDIA GeForce RTX 4090
Fine-Tuning (LoRA)NVIDIA A100 80GB SXM480 GBFP16NVIDIA H100 SXM5
Pre-TrainingNVIDIA B200 Blackwell192 GBFP8NVIDIA B200 Blackwell
Vision Model (72B)NVIDIA H200 SXM5141 GBFP8NVIDIA B200 Blackwell

Compare for a Workload

Configure a workload to see how each selected GPU performs. Calculations are deterministic estimates based on architectural specifications.

Configure workload parameters above and click "Calculate VRAM Requirement" to see results.

Loading comparison engine...

All GPU Comparisons

Browse all verified head-to-head comparisons between cloud providers, hardware configurations, and LLM APIs.