You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Short answer: 15 of them, and none of the 15 asks for a card. They are the "provider free tier" rows in this repo — permanent tiers, not trial credits that expire. Every limit below is copied from the provider's own documentation page, and the whole set was re-read on 2026-07-25.
The ones with a published number
Provider
Published free limit
Card
Source
SiliconFlow
1,000 RPM (free models are fixed; limits are per account, not per key, and each model is counted separately)
Google Gemini API, Z.AI (GLM-4.7-Flash and GLM-4.5-Flash are listed at $0 for input, cached input and output), Mistral, Alibaba Cloud Model Studio, Moonshot (Kimi), Novita, Pollinations.AI, Ollama Cloud.
These are genuinely free to call. They just don't put a per-minute or per-day figure on a public page — Gemini, for instance, publishes the tier structure and then sends you to AI Studio for the numbers actually active on your project. So this repo records "dynamic / model-dependent" rather than inventing a figure.
What this list is not
Not a key giveaway. Nothing here distributes anyone's API key. Every row links to the provider's own signup page and you make your own.
Not "free forever" guaranteed. GitHub Models is in the dataset as a retiring tier — it shut to new users on 2026-07-30. That row is kept precisely so you can see that this happens.
Not the same as cheap enough to run. Trial credits (Cerebras, Vercel AI Gateway, IBM watsonx.ai) and metered services (Together, Nebius, Perplexity, DeepInfra, Chutes) are in the dataset but deliberately kept out of the 15 — they are a different thing and get their own table.
One thing that would help: if a limit above is already stale on your account — you hit a wall the table doesn't predict, or you got more than it says — reply with the provider and what you actually saw. A limit that moved is the single hardest thing to catch from the outside, and one report is enough to send me back to the source page.
reacted with thumbs up emoji reacted with thumbs down emoji reacted with laugh emoji reacted with hooray emoji reacted with confused emoji reacted with heart emoji reacted with rocket emoji reacted with eyes emoji
Uh oh!
There was an error while loading. Please reload this page.
Short answer: 15 of them, and none of the 15 asks for a card. They are the "provider free tier" rows in this repo — permanent tiers, not trial credits that expire. Every limit below is copied from the provider's own documentation page, and the whole set was re-read on 2026-07-25.
The ones with a published number
llama-3.3-70b-versatile;llama-3.1-8b-instantgets the same 30 RPM but 14,400 per dayThe ones that are free but publish no number
Google Gemini API, Z.AI (GLM-4.7-Flash and GLM-4.5-Flash are listed at $0 for input, cached input and output), Mistral, Alibaba Cloud Model Studio, Moonshot (Kimi), Novita, Pollinations.AI, Ollama Cloud.
These are genuinely free to call. They just don't put a per-minute or per-day figure on a public page — Gemini, for instance, publishes the tier structure and then sends you to AI Studio for the numbers actually active on your project. So this repo records "dynamic / model-dependent" rather than inventing a figure.
What this list is not
Full filterable directory, all 26 entries: https://xyzs996.github.io/free-llm-api/ · how each claim was checked: https://xyzs996.github.io/free-llm-api/methodology.html
One thing that would help: if a limit above is already stale on your account — you hit a wall the table doesn't predict, or you got more than it says — reply with the provider and what you actually saw. A limit that moved is the single hardest thing to catch from the outside, and one report is enough to send me back to the source page.
All reactions