cd /news/large-language-models/i-found-35-text-generation-llm-provi… · home topics large-language-models article
[ARTICLE · art-131397] src=dev.to ↗ pub= topic=large-language-models verified=true sentiment=↑ positive

I Found 35+ Text-Generation LLM Providers That Are Entirely Free And Replenishable

A developer compiled a curated list of more than 35 text-generation LLM API providers that offer free, self-replenishing quotas, cataloging rate limits and daily token allowances for 47 services including Google Gemini, Groq, Mistral AI, Cloudflare Workers AI, and NVIDIA NIM. The list, published on GitHub as awesome-free-byok-models, was assembled by signing up as an unpaid user and stress-testing each provider's model endpoints against its stated rate limits, with the goal of finding APIs that refresh their free tiers rather than relying on one-time credits.

by read4 min views3 publishedSep 16, 2026

I made a list of Text‑generation LLM API models that are currently available as free and self‑replenishing from their API providers, stress‑tested for coding and for everyday chat.

Below is a table for the list of LLM Providers that we currently have (47).

LLM Provider Free Tier Quota
AION Labs 15 RPM / 20,000 TPD
Agnes AI 20 RPM / 1,000 RPD
AnyAPI AI 100,000 tokens/day / No Credit Card
Auriko 500 RPM (BYOK) / 1,000 RPM (Platform) / 1,000,000 tokens/month (BYOK) / Has Permanently Free models
BazaarLink 10 RPM / 50 RPD / Free Models only
Cloudflare Workers AI 150 to 1,500 RPM / 100,000 RPD / 13,000 TPD
Cohere AI 20 RPM / 1,000 API calls per month
ElectronHub 5 RPM / $0.25 Weekly Credits
EvolveX 5 RPM / No Credit Card
FastRouter 10 RPD per model / No Billing Credits Required
Free.ai 10 RPM / 30,000 TPD / 1,000 Requests per month / Currently available self-hosted models only
FreeInference $20 CPD / 2 Max Concurrent Requests
Google Gemini 5-20 RPM / 20-500 RPD / 1M TPM / Uncapped TPD
Gonka Broker 6 RPM / ~1M tokens per month
Groq API 30 RPM / 14,400 RPD / 18,000 TPM
HelixMind 3 RPM / 50 RPD
Hugging Face Inference API $0.10/month credits (~650K tokens)
Intern AI 30 RPM / 300,000 TPM / 90,000,000 Tokens per month (3,000,000 TPD)
Kilo Gateway 5 RPM / 200 RPD
LLM.Kiwi 40 RPH / No Credit Card
LLM7.IO 40 RPM / 2,400 RPD / 128,000 Characters per Request / 1,000,000 TPD
LiteRouter Unlimited Requests (for some Free models) / 1 concurrent request / 7s Cooldown
MegaNova AI 60 RPM / 550 RPD / 200,000 TPM
Mistral AI ~2–30 RPM / 50,000 TPM shared pool
Mixlayer 20 RPM / Can be rate-limited (daily usage)
Naga AI 10 RPM / 100 RPD
NVIDIA NIM 40 RPM / Uncapped TPD
Odirouter 5 RPM / 50 RPD / Free Models Only / 2 Parallel Multimodal Queries
Ollama Cloud 1 Instance / 5-Hour Session Usage / 7-day Weekly Usage
OpenCode Zen 30 RPM / 500 RPD / 1,000,000 TPD / Daily Limits
OpenRouter 20 RPM / 50 RPD
Orcarouter Unspecified rate limits / Free models only
Poixe AI 10,000 RPD / 10,000,000 TPD
Pooled AI 1M TPD / Minimax models only
Poolside 20 RPM / 200 RPD / 150,000 TPM / 1,000,000 TPD
Requesty 200 RPD / Free Models only
Routeway AI 5 RPM / 200 RPD / 300,000 TPD
SEA-LION 10 RPM
Tokeness Unspecified Rate Limits / Free models only
TokenReply 3 RPM / Free Models Only
Void AI 100 RPM / 125,000 Daily Credits
VSLLM Unspecified Rate Limits / Free Models only
xKiro AI 5M TPD / Free models only
Yolo-Auto 15 RPD
Z.AI (Zhipu AI) 1 Concurrent Request / Uncapped TPD
Zydit AI Unlimited Requests / 10 RPM / Free models only (For v3 endpoints)
Zylo API 10 RPM / 7,200 RPD / 200,000 TPD

⭐Explore the entire curated list and star the repository on GitHub I wished to have a LLM API that is needed for my long-term projects. There are so many options I can choose from, but sometimes it always fails in the long run.

The main point of failure are the free trials and one-time credits, that once drained can never be refilled again.

I promised for Free LLM APIs to follow these two main goals:

I had a third goal if second condition fails, that is:

So I've gone throughout the web, sailing across websites, as I went through several popular AI agents or libraries with their provider documentation showing a list of supported providers.

The goal is to sign myself up as a normal, unpaid user whose requirement is to know whether it refreshes its free quota or not.

I then stress-tested their model endpoints from each provider according to their rate limits.

The number of LLM Providers has dropped unexpectedly due to not having a single model found to be 200 OK but requiring payment.

To store it as helpful information, I made a repository that showcases all verified providers and its models under a single README.

Hence, I named it "Free BYOK Models", where it looks for "** Free AI models from their BYOK-compatible LLM Providers that are free and replenishable.**"

The repository is updated regularly to make it up-to-date with its latest model lists.

If you find a provider that is free and replenishable, and has no gated services, feel free to comment, or submit a PR since it updates fast.

── more in #large-language-models 4 stories · sorted by recency
── more on @google gemini 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/i-found-35-text-gene…] indexed:0 read:4min 2026-09-16 ·