⚡Under $0.50/hr🧠VRAM Estimator⚖Compare GPUs🎁Free LLM APIs🎯Model Index
Provider Directory

Free LLM Providers

Browse all verified free LLM providers. No credit card required. Compare modalities, rate limits, and hardware requirements.

Showing 14 of 14 providers

G

Google AI Studio

No Card

Google's official AI studio offering Gemini 2.0 Flash, Gemini 2.5 Flash, and Gemma 2 models with generous free quotas.

TextReasoningVisionCodeAudioEmbedding
15 RPM, 1M tokens/daySign Up →
G

Groq

No Card

Ultra-low latency inference via custom LPU chips. Fastest free tier at 330+ tok/s on Llama 3.3 70B.

TextCodeReasoning
30 RPM, 14,400 req/daySign Up →
S

SambaNova Systems

No Card

Reconfigurable dataflow unit (RDU) inference cloud. Sub-second latency on SN40L RDUs with daily free quota.

TextReasoningCode
~20 RPM, 200 RPD daily quotaSign Up →
C

Cerebras

No Card

Wafer-scale AI inference on custom CS-3 systems. Fastest free tier at 2,100+ tok/s on Llama 3.3 70B.

TextCodeReasoning
30 RPM, 1M tokens/daySign Up →
N

NVIDIA NIM

Free Credits

NVIDIA inference microservices with 1,000 free promotional API credits. Enterprise-grade GPU-accelerated inference.

TextReasoningVisionCode
1,000 promotional API credits for developersSign Up →
M

Mistral AI

Phone Verify

European open-source LLM platform with MoE architecture. Free tier requires phone verification.

TextReasoningCode
1 RPS (60 RPM), phone verification requiredSign Up →
C

Cloudflare Workers AI

No Card

Edge AI inference on Cloudflare's global network. 10,000 neurons/day free for text, audio, vision, and embeddings.

TextReasoningVisionAudioEmbedding
10,000 neurons/day free allocationSign Up →
O

OpenRouter

Free Models

Multi-provider routing layer aggregating 35+ free model endpoints. :free models available with rate limits.

TextReasoningCodeVision
3 RPM, no credit card required for :free modelsSign Up →
H

Hugging Face

Community

Open-source model hub with serverless Inference API. 1,000 requests/day free on serverless endpoints.

TextCodeEmbedding
1,000 requests/day via Inference API (serverless)Sign Up →
G

GitHub Models

Free Credits

Free AI models via GitHub Models with GitHub account. GPT-4o mini, Llama 3.3, Phi-4 available with Copilot rate limits.

TextCodeReasoning
GitHub account required; Copilot limits applySign Up →
K

Kilo Code / Kilo Gateway

Free Models

Free AI model gateway routing to frontier models including Qwen 2.5 Coder, DeepSeek Coder, and MiMo V2.5 with zero-cost access.

TextReasoningCode
Developer trial quota; varies by modelSign Up →
C

Chutes.ai

Community

Community-hosted inference with free tiers. Fast deployment for open-source LLMs with competitive rates.

TextReasoningCode
Free tier with rate limitsSign Up →
M

ModelScope

Community

Chinese AI model hub with free inference endpoints for text, code, and vision models.

TextCodeVision
Free tier with rate limitsSign Up →
O

OVHcloud AI Endpoints

Community

European sovereign AI endpoints with free tier. GDPR-compliant inference on OVHcloud infrastructure.

TextCode
Free tier with rate limitsSign Up →