⚡Under $0.50/hr🧠VRAM Estimator⚖Compare GPUs🎁Free LLM APIs🎯Model Index
MultimodalMoEContext: 125K

Ling 3.0 Flash VL (Free)

Comprehensive deployment profile and benchmark telemetry for Ling 3.0 Flash VL (Free). Self-hosting memory footprints, verified token economics, and direct API endpoints.

Verified Engineering Benchmarks

SWE-bench
49.2%
LiveCodeBench
65.9%
MATH-500
97.3%
MMLU
90.8%

VRAM Requirements & Sizing

FP16 Weights
—
INT4 / GGUF (Quantized)
—
Recommended GPU
Kilo Gateway
🔗 Quick Actions
🖥️ Compatible GPUs & Self-Host Pricing

Deterministic VRAM math from entity graph. Green = fits in single GPU.

GPUVRAMFP16FP8INT4Calculator Link
H100 SXM580 GBOOM✓ fitsOOMPre-filled →
H200141 GBOOM✓ fitsOOMPre-filled →
B200192 GBOOM✓ fitsOOMPre-filled →
A100 80GB80 GBOOM✓ fitsOOMPre-filled →
L40S48 GBOOM✓ fitsOOMPre-filled →
RTX 409024 GBOOM✓ fitsOOMPre-filled →

Active API Providers & Free Tiers

ProviderInput / 1MOutput / 1MSpeedFree Tier Status
Kilo Gateway$0.00$0.00180 tok/s⚡ Free in Kilo Code

What should I do next?