| Type | Ongoing free tier · renewing-quota | Ongoing free tier · renewing-quota |
| What's free | 10,000 Neurons/day, all account plans | Gemini 2.5 Flash, 2.5 Flash-Lite, 2.5 Pro (limited), embeddings, TTS models |
| Rate limits | 10,000 Neurons/day (free allocation). Per-task request limits: text generation 300 requests/min (models that require Workers Paid: 20 requests/min per model); text embeddings 3,000/min; speech recognition, text-to-image, translation, image-to-text 720/min | Per-model; Google no longer publishes the free-tier numbers in its public docs — they are shown only on the AI Studio rate-limit page after sign-in |
| Expires | no expiry | no expiry |
| The catch | Resets daily at 00:00 UTC; overage on a Workers Paid plan bills at $0.011/1,000 Neurons. A few models (e.g. Kimi K2.6/K2.7-code, GLM-5.2) now require a Workers Paid plan | Free-tier prompts/outputs may be used by Google to improve its products outside the UK/CH/EEA/EU. Since the 2026-03-23 terms, only Paid Services may serve API clients to end users in the EEA/CH/UK |
| Credit card | not required | not required |
| Phone verification | not required | not required |
| Commercial use | allowed | allowed |
| OpenAI-compatible | yes https://api.cloudflare.com/client/v4/accounts/{account_id}/ai/v1
| yes https://generativelanguage.googleapis.com/v1beta/openai/
|
| Modalities | text, embeddings, image, audio | text, vision, embeddings, audio |
| Free models (sample) | @cf/meta/llama-3.3-70b-instruct-fp8-fast @cf/meta/llama-3.1-8b-instruct @cf/mistralai/mistral-small-3.1-24b-instruct @cf/qwen/qwen2.5-coder-32b-instruct @cf/deepseek-ai/deepseek-r1-distill-qwen-32b @cf/google/gemma-3-12b-it | gemini-2.5-flash gemini-2.5-flash-lite gemini-2.5-pro |
| Official docs | developers.cloudflare.com/workers-ai/platform/pricing/ | ai.google.dev/gemini-api/docs/rate-limits |
| Last verified | 2026-10-08 | 2026-10-08 |