# I Found 35+ Text-Generation LLM Providers That Are Entirely Free And Replenishable

> Source: <https://dev.to/velo4705/i-found-35-text-generation-llm-providers-that-are-entirely-free-and-replenishable-226e>
> Published: 2026-09-16 12:13:58+00:00

**I made a list of Text‑generation LLM API models that are currently available as free and self‑replenishing from their API providers, stress‑tested for coding and for everyday chat.**

Below is a table for the **list of LLM Providers** that we currently have (47).

| LLM Provider | Free Tier Quota | 
|---|---|
| AION Labs | 15 RPM / 20,000 TPD | 
| Agnes AI | 20 RPM / 1,000 RPD | 
| AnyAPI AI | 100,000 tokens/day / No Credit Card | 
| Auriko | 500 RPM (BYOK) / 1,000 RPM (Platform) / 1,000,000 tokens/month (BYOK) / Has Permanently Free models | 
| BazaarLink | 10 RPM / 50 RPD / Free Models only | 
| Cloudflare Workers AI | 150 to 1,500 RPM / 100,000 RPD / 13,000 TPD | 
| Cohere AI | 20 RPM / 1,000 API calls per month | 
| ElectronHub | 5 RPM / $0.25 Weekly Credits | 
| EvolveX | 5 RPM / No Credit Card | 
| FastRouter | 10 RPD per model / No Billing Credits Required | 
| Free.ai | 10 RPM / 30,000 TPD / 1,000 Requests per month / Currently available self-hosted models only | 
| FreeInference | $20 CPD / 2 Max Concurrent Requests | 
| Google Gemini | 5-20 RPM / 20-500 RPD / 1M TPM / Uncapped TPD | 
| Gonka Broker | 6 RPM / ~1M tokens per month | 
| Groq API | 30 RPM / 14,400 RPD / 18,000 TPM | 
| HelixMind | 3 RPM / 50 RPD | 
| Hugging Face Inference API | $0.10/month credits (~650K tokens) | 
| Intern AI | 30 RPM / 300,000 TPM / 90,000,000 Tokens per month (3,000,000 TPD) | 
| Kilo Gateway | 5 RPM / 200 RPD | 
| LLM.Kiwi | 40 RPH / No Credit Card | 
| LLM7.IO | 40 RPM / 2,400 RPD / 128,000 Characters per Request / 1,000,000 TPD | 
| LiteRouter | Unlimited Requests (for some Free models) / 1 concurrent request / 7s Cooldown | 
| MegaNova AI | 60 RPM / 550 RPD / 200,000 TPM | 
| Mistral AI | ~2–30 RPM / 50,000 TPM shared pool | 
| Mixlayer | 20 RPM / Can be rate-limited (daily usage) | 
| Naga AI | 10 RPM / 100 RPD | 
| NVIDIA NIM | 40 RPM / Uncapped TPD | 
| Odirouter | 5 RPM / 50 RPD / Free Models Only / 2 Parallel Multimodal Queries | 
| Ollama Cloud | 1 Instance / 5-Hour Session Usage / 7-day Weekly Usage | 
| OpenCode Zen | 30 RPM / 500 RPD / 1,000,000 TPD / Daily Limits | 
| OpenRouter | 20 RPM / 50 RPD | 
| Orcarouter | Unspecified rate limits / Free models only | 
| Poixe AI | 10,000 RPD / 10,000,000 TPD | 
| Pooled AI | 1M TPD / Minimax models only | 
| Poolside | 20 RPM / 200 RPD / 150,000 TPM / 1,000,000 TPD | 
| Requesty | 200 RPD / Free Models only | 
| Routeway AI | 5 RPM / 200 RPD / 300,000 TPD | 
| SEA-LION | 10 RPM | 
| Tokeness | Unspecified Rate Limits / Free models only | 
| TokenReply | 3 RPM / Free Models Only | 
| Void AI | 100 RPM / 125,000 Daily Credits | 
| VSLLM | Unspecified Rate Limits / Free Models only | 
| xKiro AI | 5M TPD / Free models only | 
| Yolo-Auto | 15 RPD | 
| Z.AI (Zhipu AI) | 1 Concurrent Request / Uncapped TPD | 
| Zydit AI | Unlimited Requests / 10 RPM / Free models only (For v3 endpoints) | 
| Zylo API | 10 RPM / 7,200 RPD / 200,000 TPD | 

[⭐Explore the entire curated list and star the repository on GitHub](https://github.com/velo4705/awesome-free-byok-models)

I wished to have a LLM API that is needed for my long-term projects. There are so many options I can choose from, but sometimes it always fails in the long run.

The main point of failure are the **free trials** and **one-time credits**, that once drained can never be refilled again.

I promised for Free LLM APIs to follow these two main goals:

I had a third goal if second condition fails, that is:

So I've gone throughout the web, sailing across websites, as I went through several popular AI agents or libraries with their provider documentation showing a list of supported providers.

The goal is to **sign myself up** as a normal, unpaid user whose requirement is to know whether it refreshes its free quota or not. 

I then stress-tested their **model endpoints** from each provider according to their rate limits.

**The number of LLM Providers has dropped** unexpectedly due to not having a single model found to be `200 OK` but requiring payment.

To store it as helpful information, I made a repository that **showcases all verified providers and its models** under a single README.

Hence, I named it "**Free BYOK Models**", where it looks for "** Free AI models from their BYOK-compatible LLM Providers that are free and replenishable.**" 

The repository is **updated regularly** to make it up-to-date with its latest model lists. 

If you find a provider that is **free and replenishable, and has no gated services**, feel free to comment, or submit a PR since it updates fast.
