What people actually run for free
We pulled real usage from the gateway — every generation served for free from launch through July 11, 2026 — and ranked the underlying models. The shape is lopsided: two models carry most of the free traffic while a long tail each sees a handful of calls.
Most-used free models
Free-served requests, all-time through July 11, 2026
GLM-5.2 dominates by both requests and tokens (north of 200M free tokens served) because it's a strong, fast coding model with enough free capacity to run real work on. gpt-4o-mini is the opposite: many tiny requests, the classifier-and-glue model inside loops. The full breakdown is in /blog/most-used-free-models.
How is any of this free? Three mechanisms
"Free" here isn't a marketing asterisk — it's three concrete routes, each with clear limits:
- The anyrouter/free tier — a rotating router over whatever free capacity is healthy. Requires the Go plan ($2/mo, or donate a key); capped at 1,000 requests/day, no credit spent.
- The donated shared pool — members donate spare provider quota; those keys pool together and members draw on the sum. BYOK traffic never deducts credits, and donors earn 8% back.
- Self-hosted on-device models — a first-party local relay serves models running on your own paired hardware, including Apple's on-device Foundation Model, at no provider cost.
Each is honest about its ceiling: the daily tier has a request cap, the pool depends on donated headroom (and requires you to donate to draw on it on the Free plan), and on-device inference needs your own hardware.
Start with free models
Read /free for the anyrouter/free tier (Go plan, 1,000 req/day), browse the catalog on /models, and donate an idle key at /donate to unlock the pool and earn credits back.
- The $2 (or free) plan that unlocks the free tier — /blog/free-models-one-dollar.
- How the donated pool pays donors back — /blog/shared-key-pool.
- On-device inference through the gateway — /blog/apple-foundation-models-on-device.
A real free LLM API — the free tier, the donated pool, and on-device models in one place.
Try free modelsRelated posts
Free LLM API: 150+ models with a real free tier
You don't need a credit card to start building with large language models. A free LLM API should give you real models, an OpenAI-compatible endpoint, and a way to keep costs at zero as you grow — here's how that works in practice.
The most-used free models on AnyRouter so far
We pulled real usage from the gateway: which free models people actually route to. GLM-5.2 runs away with it, gpt-4o-mini is the workhorse behind it, and anyrouter/free quietly serves the long tail. Numbers below, straight from production.
Route your first request in 2 minutes
Start free with your own keys, or top up and pay per token. Get $4/mo in credits and free models on Go — $2/mo, or free when you donate a provider key.
Start free