{"slug": "editorial-re-verification-freeinference", "title": "Editorial re-verification: FreeInference", "summary": "FreeInference, an OpenAI- and Anthropic-compatible inference API built at Harvard SEAS MadSys Lab and sponsored by NVIDIA and Harvard SEAS, paused new account onboarding on Aug 11, 2026 because the service is at capacity. The free research service publishes no numeric quota — its /pricing and /limits paths both return 404 — and warns that GLM, Minimax and Kimi models will be discontinued on Oct 1. Its Terms of Service allow prompts and responses to be logged for research and anonymized derived data to be published or open-sourced, ruling out confidential or regulated workloads.", "body_md": "# > FreeInference\n\n[freeinference.org](https://freeinference.org/)\n\nFreeInference (freeinference.org) is an OpenAI- and Anthropic-compatible inference API offering frontier open models free of charge to the research community, with no credit card required at signup. It is presented as built at Harvard SEAS (MadSys Lab) and lists NVIDIA and Harvard SEAS as sponsors. The quota is described only as a \"generous quota for research and prototyping\"; no numeric limit is published, and the /pricing and /limits paths both return 404. The site carries a banner warning that GLM, Minimax and Kimi will be discontinued on Oct 1, and on Aug 11, 2026 it announced that new account onboarding is paused because the service is at capacity.\n\n### // pros\n\n- Drop-in OpenAI and Anthropic compatible endpoints: the documented base URL is https://freeinference.org/v1/chat/completions, so existing SDK clients work unchanged.\n- Free to use with no credit card required at signup, positioned for research, education and prototyping.\n- Research-lab backing: the site credits Harvard SEAS MadSys Lab and displays NVIDIA and Harvard SEAS as sponsors.\n- Coding-agent oriented use cases are named explicitly (Claude Code, Kilo, Hermes, OpenClaw) rather than only generic chat.\n- A public curl quickstart is published on the homepage, so the API shape can be evaluated before an account exists.\n\n### // cons\n\n- New signups are closed: the Aug 11, 2026 site notice states onboarding is stopped because the service is at capacity, so a new user cannot obtain a key today.\n- The Terms of Service state that all prompts and responses may be logged for research purposes and that anonymized derived data may be published or open-sourced, which rules out confidential or regulated workloads.\n- Model availability is unstable: GLM, Minimax and Kimi are scheduled for discontinuation on Oct 1, and quotas and model access may change without notice.\n- No quantified free tier exists on any official page: /pricing and /limits return 404, and the homepage gives only the phrase \"Generous quota\"; there is no published SLA, paid tier or upgrade path.\n- Eligibility requires the user to be at least 18 years old, and access may be limited, suspended or terminated for operational needs.\n\nUse FreeInference for research prototypes, benchmarks and agent experiments where the prompts are not sensitive and the workload can tolerate a model vanishing on short notice. Because signups are currently closed, treat it as a supplementary endpoint rather than a primary one, and keep a paid provider configured for anything that must stay reproducible. Anyone comparing free options in this category should check whether onboarding has reopened before planning work around it.\n\nimported analysis — independent third party, not Uprouter editorial · verify before relying on it\n\nFree tier present — value not quantified\n\nThis provider offers a free tier, but its terms don’t convert cleanly into a USD figure (e.g. request-count limits or capacity-dependent pools). We refuse to print a misleading zero here. See the note below and the official pricing page for the exact limits.\n\nnote: No free tier stated on official pages (checked 2026-09-16): freeinference.org offers a free account with a \"Generous quota for research and prototyping\" but publishes no dollar or token figure, and freeinference.org/pricing returns 404.\n\n| Plan | Type | Monthly | Input /1M | Output /1M | Note | \n|---|---|---|---|---|---|\n| Pay-as-you-go | payg | — | Free | Free | Cheapest tracked model rate (OpenRouter-normalized). | \n| Research Free Tier | free | — | — | — | Generous quota for research and prototyping; no credit card required. | \n\n[official pricing page](https://freeinference.org/)\n\nPricing normalized from public sources — always verify with the provider.\n\n### // model prices\n\n| Model | Input /1M | Output /1M | Context | \n|---|---|---|---|\n| [Qwen3 Coder](https://www.uprouter.online/models/qwen3-coder) | Free | $0.0000 | 262k | \n| [MiniMax-M3](https://www.uprouter.online/models/minimax-m3) | Free | $0.0000 | 1.0M | \n| [DeepSeek V4 Flash](https://www.uprouter.online/models/deepseek-v4-flash) | Free | Free | 1.0M | \n| [GLM-5.1](https://www.uprouter.online/models/glm-5-1) | $0.0000 | $0.0000 | 205k | \n\n## $ Does FreeInference have a free tier?\n\nYes — FreeInference offers a free tier. Always confirm current limits on FreeInference's official pricing page — free tiers change without notice.\n\n## $ How much does FreeInference cost per 1M tokens?\n\nThe cheapest model we track at FreeInference is Free per 1M input tokens and Free per 1M output tokens. This is normalized from FreeInference's published pricing — verify with the provider before purchasing, since prices change frequently.\n\n## $ Is FreeInference safe to use?\n\nUprouter rates FreeInference as low risk. Risk is high, driven mainly by service stability rather than by data policy alone. The operator is a university research lab project (Harvard SEAS MadSys Lab) rather than a commercial vendor, its Terms describe it as an experimental research service provided as-is with no performance guarantee, and it has already stopped onboarding with a capacity notice. The Terms explicitly permit logging of prompts and responses for research and allow de-identified derived data to be published or open-sourced, with sanitization acknowledged as not guaranteed. There is no published business model, paid tier or funding disclosure, so the service can be discontinued at any time without a commercial consequence for the operator. Risk ratings are editorial, evidence-based, and never influenced by affiliate relationships.\n\n## $ Can I use FreeInference through Uprouter Connect?\n\nYes — FreeInference is Connect-compatible. Add your FreeInference API key in Uprouter Connect (encrypted at rest) and route requests to it through one OpenAI- and Claude-compatible API, optionally behind failover aliases. Uprouter meters the traffic in compute at the provider's listed rate.\n\nRaw community sentiment, shown unblended. Votes never feed into objective data fields like price, risk, or status — and never into sort order.\n\nNo published reviews yet — be the first.\n\n## // more_informationprice history · live status · risk & confidence · code snippets\n\n### price_history\n\nNo recorded changes yet.\n\n### live_status\n\ncurrent: unknown\nNo probes recorded yet.\n\nUprouter runs its own probes; best-effort, not a guarantee. Probe results depend on our network location and cadence — for production, always use the provider’s own status page.\n\n### risk_&_confidence\n\nRisk is high, driven mainly by service stability rather than by data policy alone. The operator is a university research lab project (Harvard SEAS MadSys Lab) rather than a commercial vendor, its Terms describe it as an experimental research service provided as-is with no performance guarantee, and it has already stopped onboarding with a capacity notice. The Terms explicitly permit logging of prompts and responses for research and allow de-identified derived data to be published or open-sourced, with sanitization acknowledged as not guaranteed. There is no published business model, paid tier or funding disclosure, so the service can be discontinued at any time without a commercial consequence for the operator.\n\nConfidence reflects how much verifiable evidence (official docs, probe history, community reports) backs this entry. Risk is our editorial assessment of operator reliability and terms — not financial advice.\n\n### code_snippets\n\n```\ncurl https://www.uprouter.online/api/connect/v1/chat/completions \\  -H \"Authorization: Bearer upr_live_YOUR_KEY\" \\  -H \"Content-Type: application/json\" \\  -d '{\n    \"model\": \"glm-5-1\",\n    \"messages\": [{ \"role\": \"user\", \"content\": \"Hello via FreeInference\" }]\n  }'\n```\n\nMetered in Uprouter compute — the usage log shows which upstream served each request and its exact cost.\n\n[NLP Cloudfree $15.00 est.pricing n/aconnectLow risk](https://www.uprouter.online/s/nlpcloud-com)\n\n[Nomicfree $0.10 est.pricing n/aconnectLow risk](https://www.uprouter.online/s/nomic-ai)\n\n[Lightning AIfree $15.00 est.pricing n/aconnectLow risk](https://www.uprouter.online/s/lightning-ai)\n\n[AI21 Labsfree $10.00pricing n/aconnectLow risk](https://www.uprouter.online/s/api-ai21-com)\n\nRanked by neutral similarity signals only — live status, Connect compatibility, free-tier shape and free-credit proximity. Commercial relationships never influence this list.", "url": "https://wpnews.pro/news/editorial-re-verification-freeinference", "canonical_source": "https://www.uprouter.online/s/freeinference", "published_at": "2026-09-16 21:08:07+00:00", "updated_at": "2026-09-16 21:24:28.736163+00:00", "lang": "en", "topics": ["ai-infrastructure", "ai-tools", "ai-products", "large-language-models", "ai-research"], "entities": ["FreeInference", "Harvard SEAS", "MadSys Lab", "NVIDIA", "GLM", "Minimax", "Kimi", "Qwen3 Coder"], "alternates": {"html": "https://wpnews.pro/news/editorial-re-verification-freeinference", "markdown": "https://wpnews.pro/news/editorial-re-verification-freeinference.md", "text": "https://wpnews.pro/news/editorial-re-verification-freeinference.txt", "jsonld": "https://wpnews.pro/news/editorial-re-verification-freeinference.jsonld"}}