I Found 35+ Text-Generation LLM Providers That Are Entirely Free And Replenishable A developer compiled a curated list of more than 35 text-generation LLM API providers that offer free, self-replenishing quotas, cataloging rate limits and daily token allowances for 47 services including Google Gemini, Groq, Mistral AI, Cloudflare Workers AI, and NVIDIA NIM. The list, published on GitHub as awesome-free-byok-models, was assembled by signing up as an unpaid user and stress-testing each provider's model endpoints against its stated rate limits, with the goal of finding APIs that refresh their free tiers rather than relying on one-time credits. I made a list of Text‑generation LLM API models that are currently available as free and self‑replenishing from their API providers, stress‑tested for coding and for everyday chat. Below is a table for the list of LLM Providers that we currently have 47 . | LLM Provider | Free Tier Quota | |---|---| | AION Labs | 15 RPM / 20,000 TPD | | Agnes AI | 20 RPM / 1,000 RPD | | AnyAPI AI | 100,000 tokens/day / No Credit Card | | Auriko | 500 RPM BYOK / 1,000 RPM Platform / 1,000,000 tokens/month BYOK / Has Permanently Free models | | BazaarLink | 10 RPM / 50 RPD / Free Models only | | Cloudflare Workers AI | 150 to 1,500 RPM / 100,000 RPD / 13,000 TPD | | Cohere AI | 20 RPM / 1,000 API calls per month | | ElectronHub | 5 RPM / $0.25 Weekly Credits | | EvolveX | 5 RPM / No Credit Card | | FastRouter | 10 RPD per model / No Billing Credits Required | | Free.ai | 10 RPM / 30,000 TPD / 1,000 Requests per month / Currently available self-hosted models only | | FreeInference | $20 CPD / 2 Max Concurrent Requests | | Google Gemini | 5-20 RPM / 20-500 RPD / 1M TPM / Uncapped TPD | | Gonka Broker | 6 RPM / ~1M tokens per month | | Groq API | 30 RPM / 14,400 RPD / 18,000 TPM | | HelixMind | 3 RPM / 50 RPD | | Hugging Face Inference API | $0.10/month credits ~650K tokens | | Intern AI | 30 RPM / 300,000 TPM / 90,000,000 Tokens per month 3,000,000 TPD | | Kilo Gateway | 5 RPM / 200 RPD | | LLM.Kiwi | 40 RPH / No Credit Card | | LLM7.IO | 40 RPM / 2,400 RPD / 128,000 Characters per Request / 1,000,000 TPD | | LiteRouter | Unlimited Requests for some Free models / 1 concurrent request / 7s Cooldown | | MegaNova AI | 60 RPM / 550 RPD / 200,000 TPM | | Mistral AI | ~2–30 RPM / 50,000 TPM shared pool | | Mixlayer | 20 RPM / Can be rate-limited daily usage | | Naga AI | 10 RPM / 100 RPD | | NVIDIA NIM | 40 RPM / Uncapped TPD | | Odirouter | 5 RPM / 50 RPD / Free Models Only / 2 Parallel Multimodal Queries | | Ollama Cloud | 1 Instance / 5-Hour Session Usage / 7-day Weekly Usage | | OpenCode Zen | 30 RPM / 500 RPD / 1,000,000 TPD / Daily Limits | | OpenRouter | 20 RPM / 50 RPD | | Orcarouter | Unspecified rate limits / Free models only | | Poixe AI | 10,000 RPD / 10,000,000 TPD | | Pooled AI | 1M TPD / Minimax models only | | Poolside | 20 RPM / 200 RPD / 150,000 TPM / 1,000,000 TPD | | Requesty | 200 RPD / Free Models only | | Routeway AI | 5 RPM / 200 RPD / 300,000 TPD | | SEA-LION | 10 RPM | | Tokeness | Unspecified Rate Limits / Free models only | | TokenReply | 3 RPM / Free Models Only | | Void AI | 100 RPM / 125,000 Daily Credits | | VSLLM | Unspecified Rate Limits / Free Models only | | xKiro AI | 5M TPD / Free models only | | Yolo-Auto | 15 RPD | | Z.AI Zhipu AI | 1 Concurrent Request / Uncapped TPD | | Zydit AI | Unlimited Requests / 10 RPM / Free models only For v3 endpoints | | Zylo API | 10 RPM / 7,200 RPD / 200,000 TPD | ⭐Explore the entire curated list and star the repository on GitHub https://github.com/velo4705/awesome-free-byok-models I wished to have a LLM API that is needed for my long-term projects. There are so many options I can choose from, but sometimes it always fails in the long run. The main point of failure are the free trials and one-time credits , that once drained can never be refilled again. I promised for Free LLM APIs to follow these two main goals: I had a third goal if second condition fails, that is: So I've gone throughout the web, sailing across websites, as I went through several popular AI agents or libraries with their provider documentation showing a list of supported providers. The goal is to sign myself up as a normal, unpaid user whose requirement is to know whether it refreshes its free quota or not. I then stress-tested their model endpoints from each provider according to their rate limits. The number of LLM Providers has dropped unexpectedly due to not having a single model found to be 200 OK but requiring payment. To store it as helpful information, I made a repository that showcases all verified providers and its models under a single README. Hence, I named it " Free BYOK Models ", where it looks for " Free AI models from their BYOK-compatible LLM Providers that are free and replenishable. " The repository is updated regularly to make it up-to-date with its latest model lists. If you find a provider that is free and replenishable, and has no gated services , feel free to comment, or submit a PR since it updates fast.