{"slug": "i-tested-100-providers-to-find-free-and-replenishable-text-generation-llm-apis", "title": "I Tested 100+ Providers to Find Free and Replenishable Text-generation LLM APIs", "summary": "A developer tested over 100 LLM providers to find free and replenishable text-generation APIs, compiling verified results into a GitHub repository named 'Free BYOK Models'. The list includes providers like Google Gemini, Groq, and Cloudflare Workers AI, with details on their free tier quotas.", "body_md": "I searched around **hundreds of LLM Providers** across the web and GitHub, To find LLM APIs that have a free tier and should automatically refill itself on a cycle, whether it could be daily, weekly or monthly.\n\nThis all started back where I wish to have a LLM API that is needed for my long-term projects. There are so many options I can choose from, but sometimes I just needed a lot of \"fallbacks\" to support my project. \n\nBut to prevent it from failing, the LLM APIs are needed to **replenish itself**, which considering I found a few of them and they are extremely common, It felt like it was not enough. \n\nHaving to add some providers which would have a one-time credit or a free trial would really not fit it, considering it would be drained in the long run, and never be able to use it again.\n\nSo, I've gone throughout the web, sailing across websites that promise Free LLM APIs to follow these two main goals:\n\nI also have a third goal if second condition fails: **If it's not Replenishable, it must have at least one free model available in their catalog that works without wasting a single credit dime (>=$0.0000001)**, and without asking to require credit top-ups.\n\nI went through several popular AI agents with their provider documentation showing a list of providers, so that my goal is to **sign myself up** as a normal, unpaid user whose requirement is to know whether it refreshes its free quota or not. Not only that, I did explore a lot of GitHub Repositories finding providers that curators add onto their lists.\n\nAfter I found many LLM providers that \"promises\" those two rules, it is not the end yet. So, I stress-tested their **model endpoints** from each provider according to their rate limits. \n\nSurprisingly, **the number of LLM Providers has dropped** due to not having a single model found to be `200 OK` but requiring payment. However, not all providers are dropped when some providers **have one or more models** recieving `200 OK` from just a single prompt, without spending tokens.\n\nTo store it as helpful information, I made a repository that **showcases all verified providers and its models** under a single README. The number of providers that I found across the internet seemed to increase slow and steady overtime due to providers that are seen as obscure, but follows the main two golden rules that keeps the list on-topic as it was.\n\nHence, it is named \"**Free BYOK Models**\", where the theme suggests for \"** Free AI models from their BYOK-compatible LLM Providers that are free and replenishable.**\" \n\nThis is solely for **Text-generation endpoints**, but it can support Vision-compatible models too. It is useful for trying to integrate API keys using these providers into their AI assistants, Coding agents, etc.\n\nThe following table shows **available LLM providers** along with their Free Tier Quota, at the time of writing:\n\n| LLM Provider | Free Tier Quota | \n|---|---|\n| AION Labs | 15 RPM / 20,000 TPD | \n| Agnes AI | 20 RPM / 1,000 RPD | \n| AnyAPI AI | 100,000 tokens/day / No Credit Card | \n| Auriko | 500 RPM (BYOK) / 1,000 RPM (Platform) / 1,000,000 tokens/month (BYOK) / Has Permanently Free models | \n| BazaarLink | 10 RPM / 50 RPD / Free Models only | \n| Cloudflare Workers AI | 150 to 1,500 RPM / 100,000 RPD / 13,000 TPD | \n| Cohere AI | 20 RPM / 1,000 API calls per month | \n| ElectronHub | 5 RPM / $0.25 Weekly Credits | \n| EvolveX | 5 RPM / No Credit Card | \n| FastRouter | 10 RPD per model / No Billing Credits Required | \n| Free.ai | 10 RPM / 30,000 TPD / 1,000 Requests per month / Currently available self-hosted models only | \n| FreeInference | $20 CPD / 2 Max Concurrent Requests | \n| Google Gemini | 5-20 RPM / 20-500 RPD / 1M TPM / Uncapped TPD | \n| Gonka Broker | 6 RPM / ~1M tokens per month | \n| Groq API | 30 RPM / 14,400 RPD / 18,000 TPM | \n| HelixMind | 3 RPM / 50 RPD | \n| Hugging Face Inference API | $0.10/month credits (~650K tokens) | \n| Intern AI | 30 RPM / 300,000 TPM / 90,000,000 Tokens per month (3,000,000 TPD) | \n| Kilo Gateway | 5 RPM / 200 RPD | \n| LLM.Kiwi | 40 RPH / No Credit Card | \n| LLM7.IO | 40 RPM / 2,400 RPD / 128,000 Characters per Request / 1,000,000 TPD | \n| LiteRouter | Unlimited Requests (for some Free models) / 1 concurrent request / 7s Cooldown | \n| MegaNova AI | 60 RPM / 550 RPD / 200,000 TPM | \n| Mistral AI | ~2–30 RPM / 50,000 TPM shared pool | \n| Mixlayer | 20 RPM / Can be rate-limited (daily usage) | \n| Naga AI | 10 RPM / 100 RPD | \n| NVIDIA NIM | 40 RPM / Uncapped TPD | \n| Odirouter | 5 RPM / 50 RPD / Free Models Only / 2 Parallel Multimodal Queries | \n| Ollama Cloud | 1 Instance / 5-Hour Session Usage / 7-day Weekly Usage | \n| OpenCode Zen | 30 RPM / 500 RPD / 1,000,000 TPD / Daily Limits | \n| OpenRouter | 20 RPM / 50 RPD | \n| Orcarouter | Unspecified rate limits / Free models only | \n| Poixe AI | 10,000 RPD / 10,000,000 TPD | \n| Pooled AI | 1M TPD / Minimax models only | \n| Poolside | 20 RPM / 200 RPD / 150,000 TPM / 1,000,000 TPD | \n| Requesty | 200 RPD / Free Models only | \n| Routeway AI | 5 RPM / 200 RPD / 300,000 TPD | \n| SEA-LION | 10 RPM | \n| TokenReply | 3 RPM / Free Models Only | \n| Void AI | 100 RPM / 125,000 Daily Credits | \n| xKiro AI | 5M TPD / Free models only | \n| Yolo-Auto | 15 RPD | \n| Z.AI (Zhipu AI) | 1 Concurrent Request / Uncapped TPD | \n| Zydit AI | Unlimited Requests / 10 RPM / Free models only (For v3 endpoints) | \n| Zylo API | 10 RPM / 7,200 RPD / 200,000 TPD | \n\nThe repository below is a curation list that contains **model lists** of verified LLM providers along with their **latencies**, **Context Windows**, **free tier quotas**, and **Base URLs**. \n\nAt the time of writing this post, there are **45 LLM Providers** that resets their quota in a time cycle.\n\nThe repository is **updated regularly** to make it up-to-date with its latest model lists. If you find a provider that is **free and replenishable, and has no gated services**, feel free to comment, or submit a PR since it updates fast.\n\n⭐ **[Star the Repository](https://github.com/velo4705/awesome-free-byok-models)** if you find this useful!", "url": "https://wpnews.pro/news/i-tested-100-providers-to-find-free-and-replenishable-text-generation-llm-apis", "canonical_source": "https://dev.to/velo4705/i-tested-100-providers-to-find-free-and-replenishable-text-generation-llm-apis-15ha", "published_at": "2026-09-09 12:12:01+00:00", "updated_at": "2026-09-09 12:41:46.401400+00:00", "lang": "en", "topics": ["large-language-models", "developer-tools", "ai-products"], "entities": ["Google Gemini", "Groq", "Cloudflare Workers AI", "Hugging Face", "Cohere", "AION Labs", "Agnes AI", "Free BYOK Models"], "alternates": {"html": "https://wpnews.pro/news/i-tested-100-providers-to-find-free-and-replenishable-text-generation-llm-apis", "markdown": "https://wpnews.pro/news/i-tested-100-providers-to-find-free-and-replenishable-text-generation-llm-apis.md", "text": "https://wpnews.pro/news/i-tested-100-providers-to-find-free-and-replenishable-text-generation-llm-apis.txt", "jsonld": "https://wpnews.pro/news/i-tested-100-providers-to-find-free-and-replenishable-text-generation-llm-apis.jsonld"}}