Google AI Studio
No CardGoogle's official AI studio offering Gemini 2.0 Flash, Gemini 2.5 Flash, and Gemma 2 models with generous free quotas.
Browse all verified free LLM providers. No credit card required. Compare modalities, rate limits, and hardware requirements.
Showing 14 of 14 providers
Google's official AI studio offering Gemini 2.0 Flash, Gemini 2.5 Flash, and Gemma 2 models with generous free quotas.
Ultra-low latency inference via custom LPU chips. Fastest free tier at 330+ tok/s on Llama 3.3 70B.
Reconfigurable dataflow unit (RDU) inference cloud. Sub-second latency on SN40L RDUs with daily free quota.
Wafer-scale AI inference on custom CS-3 systems. Fastest free tier at 2,100+ tok/s on Llama 3.3 70B.
NVIDIA inference microservices with 1,000 free promotional API credits. Enterprise-grade GPU-accelerated inference.
European open-source LLM platform with MoE architecture. Free tier requires phone verification.
Edge AI inference on Cloudflare's global network. 10,000 neurons/day free for text, audio, vision, and embeddings.
Multi-provider routing layer aggregating 35+ free model endpoints. :free models available with rate limits.
Open-source model hub with serverless Inference API. 1,000 requests/day free on serverless endpoints.
Free AI models via GitHub Models with GitHub account. GPT-4o mini, Llama 3.3, Phi-4 available with Copilot rate limits.
Free AI model gateway routing to frontier models including Qwen 2.5 Coder, DeepSeek Coder, and MiMo V2.5 with zero-cost access.
Community-hosted inference with free tiers. Fast deployment for open-source LLMs with competitive rates.
Chinese AI model hub with free inference endpoints for text, code, and vision models.
European sovereign AI endpoints with free tier. GDPR-compliant inference on OVHcloud infrastructure.