⚡Under $0.50/hr🧠VRAM Estimator⚖Compare GPUs🎁Free LLM APIs🎯Model Index
WORKLOAD TRADE-OFF MATRIX

GPU Comparison Engine

Select any two GPUs for a side-by-side evaluation of VRAM capacity, memory bandwidth, TFLOPS throughput, and live spot pricing across workload profiles.

Methodology: How GPU Comparison Works

OpenGPU Radar compares GPU instances across four dimensions: Price, Performance, Memory, and Availability. Pricing data is aggregated from Spheron, RunPod, Vast.ai, and Lambda Labs APIs, refreshed every 6 hours.

Cost per 1M Tokens = (Hourly GPU Cost) / (Tokens/Second × 1,000,000) Lower values indicate better economics. Throughput Index = Effective Tokens/Second (vLLM FP8, Llama 70B) Measured end-to-end output throughput including KV-cache retrieval. Example: H100 SXM5 at $1.89/hr delivers ~120 tok/s → $1.89 / (120 × 1M) = $0.000158 per 1M tokens

Price Data Source
Cloud Provider APIs
Refresh Rate
Every 6 hours
Baseline GPU
RTX 4090 ($0.34/hr)
Benchmarks
vLLM FP8 inference

Hardware Requirements by Workload

Minimum VRAM and GPU specs required for each workload category. Determined via entity-graph VRAM calculations.

WorkloadMin GPUMin VRAMPrecisionRecommended GPU
70B LLM InferenceH100 SXM580 GBFP8H200 or B200
30B LLM InferenceL40S48 GBFP8A100 80GB
7B LLM InferenceRTX 409024 GBFP16RTX 4090
Fine-Tuning (LoRA)A100 80GB80 GBFP16H100 × 2
Pre-TrainingB200192 GBFP8B200 × 8
Vision Model (72B)H200141 GBFP8B200
Loading comparison engine...