What's free
Lite plan: 300,000 tokens/month for foundation model inference, 20 CUH/month for ML tooling, 100 pages/month of document text extraction
Rate limits
2 inference requests per second (explicitly documented for the Lite plan)
The catch
Lite plan doesn't support fine-tuning of foundation or custom models; 1-day idle deployment timeout. Never expires or bills while inside quota, but a payment method (with a nominal ~$1 authorization hold) is required at signup
Appears in
Change history
How this free tier has changed since we started tracking it (2026-07-30) — generated from the git history of providers.json.
- 2026-07-30 Updated OpenAI compatibility
- 2026-07-30 Added to the hub