Free LLM APIs: 70 verified entries, 58 of which need no credit card, last checked 2026-09-24.
A free LLM API is the starting point for almost every agent project, and it is the category where lists rot fastest: Cerebras removed its no-card free tier in July 2026, Google moved its Pro models behind billing in April, and Half the providers quoted in older guides now require a payment method before the first request.
Every entry below states the free allowance in the vendor's own units, the rate limit, whether a credit card is required, and whether the free tier trains on your prompts. Providers that quietly stopped being free are still listed — marked as no longer free, with the replacement named — because knowing where not to go is half the value.
If you want to start in the next five minutes: Gemini gives the most daily requests without a card, Groq is the fastest, and Ollama is the only option with genuinely no limit at all.
Anonymous open-model inference on a decentralised GPU network — 100M tokens per key with no account, no card and no signup, on an OpenAI-compatible endpoint.
Free limit
100M tokens per API key
Rate limit
Not published; the 100M-token balance is the binding limit
✅ no credit cardDoes not train on prompts
The 100M allowance is per key, not per day — issuance is unlimited, so you can mint another key when one is spent. Anonymous keys come from POST /tokens. Direct curl calls frequently trip a Cloudflare challenge (403); the browser chat and the documented endpoints work. Runs on the Gonka compute network.
Free OpenAI-compatible access to Gemini, GPT-5.2, Kimi K2 and 30+ other models for Hack Clubbers — the most generous free tier aimed at teenage developers.
Free limit
Free for Hack Club members
Rate limit
Not published
✅ no credit card
Gated on a Hack Club account, which is aimed at students and teenage hackers. Not a general-purpose public tier — do not list it as one.
Gateway you can call with the literal placeholder key "unused" — no account, no email, no card, just a base URL and an OpenAI SDK.
Free limit
10 req/min anonymous; 40 req/min with a free token
Rate limit
10 RPM anonymous, 40 RPM with token
Context
1M
✅ no credit card
The widely-repeated "2 req/s, 20 RPM, 100 req/hr" figures do not match the vendor documentation: anonymous access is 10 req/min and a free token from token.llm7.io raises it to 40 req/min. Anonymous traffic only reaches the turbo model set.
EU-hosted, GDPR-friendly inference over 20+ open-weight models at $0, usable without an API key on the anonymous tier.
Free limit
Free — $0 per request on the anonymous tier
Rate limit
~2 requests/min per model anonymous
✅ no credit card
The API host is oai.endpoints.kepler.ai.cloud.ovh.net — posting to the marketing domain returns 405. Anonymous calls are accepted and then answered with "API rate limit exceeded", which is how the ~2 req/min cap presents itself; no key is required to reach that point. The published catalogue returns pricing of 0 for prompt, completion and request.
Client-side JavaScript SDK that gives a web app AI, storage, database and auth with no backend — the developer pays nothing because each signed-in user spends credits from their own Puter account.
Free limit
Free for developers; end users spend their own Puter credits
Rate limit
Per-user credit balance
✅ no credit card
Usually marketed as "free and unlimited", which is only half true. It is genuinely free for the developer, but not unlimited: once a user exhausts their Puter allocation they pay Puter directly, so an app with heavy anonymous traffic will hit a wall. The open-source developer tools that wrap it as a drop-in OpenAI server are unmaintained third-party bridges and are not listed.
Anonymous stealth model on OpenRouter with a 1M-token context window, free at $0 in and $0 out while the provider stays cloaked.
Free limit
Free ($0/$0 per million tokens) while the preview runs
Rate limit
OpenRouter free-tier limits
Context
1M
✅ no credit card
Confirmed live against OpenRouter's model API on 2026-09-24. Stealth previews are temporary by design: Ox Alpha, the previous drop, was unmasked as Z.AI's GLM-5.3-Flash and removed from the free catalogue within a week — which is exactly why it is not listed here. Needs an OpenRouter account, and prompts are handled by an anonymous provider, so keep secrets out of it.
Unified gateway over 100+ models that keeps a small set of model ids permanently free, including DeepSeek V4 Flash Vision and GLM-5.3.
Free limit
Free model ids, including deepseek/deepseek-v4-flash-vision-exp-free and z-ai/glm-5.3-free
Rate limit
Not published
✅ no credit card
The gateway as a whole is commercial — only the -free suffixed model ids cost nothing. Paid tiers advertise a 5% service-fee discount on top-up, so expect upsell copy on the pricing page.
70 free LLM API providers are listed in this catalogue as of 2026-09-24, and 58 of them can be started without a credit card. Each entry records the free limit in the provider's own units, the rate limit, and the date a human last confirmed it.
Which free LLM APIs need no credit card?
58 of the 70 providers here can be used without entering card details. The catalogue flags every provider that requires a card and can filter them out in one click, which matters because a card-required free tier is a trial with extra steps.
Are free LLM APIs actually free?
They are free up to a published quota, not unmetered. Every provider listed states a limit — requests per minute, tokens per day, or a credit balance — and this catalogue records that limit instead of repeating a marketing claim. No legitimate unmetered frontier API exists, so none is listed.
What is the catch with free LLM APIs?
Three catches recur: your prompts may be used for training, commercial use may be restricted, and a free tier can be withdrawn with little notice. All three are recorded per provider here, which is why the entries carry a training flag, a commercial-use flag and a last-verified date.