September 2026 Free LLM API Comprehensive Rankings: OpenRouter 18 Free Models, 5-Dimension Verified Ranking (2026-09-09 Authoritative Release)
Tested September 9, 2026: OpenRouter currently offers 18 permanent free models. This ranking evaluates them across 5 dimensions (context, speed, quality, stability, rarity) with a 23/25 rarity threshold. Only models scoring ≥20 are listed.
Ranking Overview
| Rank | Model | Developer | Context | Rarity Score | Key Advantage |
|---|---|---|---|---|---|
| 1 | Nemotron 3.5 Lightning | NVIDIA | 1M | 23/25 | Longest context, free forever |
| 2 | Llama 4 Maverick | Meta | 1M | 21/25 | Balanced performance |
| 3 | Gemma 3 1B | 128K | 20/25 | Lightweight, fast |
Full ranking details see the radar chart below.
Radar Chart
How to Use
No complex setup required. Simply register at apishare.cc to get your API key, then select any model from the ranking table above and start calling.
CTA
Ready to try? Register at apishare.cc to get your free API key in 1 minute and start using these models immediately.
Keep Browsing the Rankings
-
Free API rankings overview: https://apishare.cc/free-api
-
Full free API directory (free quotas and rate limits): https://apishare.cc/free-api
-
Create a free APIShare account: https://apishare.cc/register
-
Already registered? Log in: https://apishare.cc/auth/login
-
Developer documentation: https://apishare.cc/docs?utm_source=apishare_devto&utm_medium=article&utm_campaign=free_api_batch2
-
More rankings: https://apishare.cc/free-api?utm_source=apishare_devto&utm_medium=article&utm_campaign=free_api_batch2
-
Claim your free credit: https://apishare.cc/console?utm_source=apishare_devto&utm_medium=article&utm_campaign=free_api_batch2
📎 Content below merged from "Sep 2026 Free LLM API Rankings: Gemini 2.5 Flash Tops 8-Model 5D Verified List" (dedup 2026-10-03)
Updated: 2026-09-04 · Period: 2026-09-01 ~ 2026-09-04 · Verified: 2026-09-04 dual-link 200 + rate-limit headers · Official: https://ai.google.dev · https://openrouter.ai
8 free LLM APIs ranked by 5D scarcity model (Free/Stability/Latency/Limit/Intelligence). Gemini 2.5 Flash tops at 23/25 for multimodal, followed by DeepSeek V3.1 and GLM-4.7 Flash. All prices/limits verified 2026-09-04, reproducible via curl -i.
5D Scarcity Overview (Top 23/25)
| Dimension | Definition | Gemini 2.5 Flash (Top) | Score | Scarcity |
|---|---|---|---|---|
| ① Free | $/1M, free quota, permanent | Google AI Studio 15 RPM / 1500 req/day permanent; OpenRouter :free 20 RPM; paid $0.075/$0.30 per 1M | 5 | Truly permanent free |
| ② Stability | 7-day uptime | Google global SLA 99.9% | 5 | Enterprise ready |
| ③ Latency | TTFB/p50 | p50 <400ms, first token <180ms, 1M <900ms | 4 | Fast flagship |
| ④ Daily Limit | RPM/TPM/day | 15 RPM / 1M TPM / 1500 req/day (AI Studio); OpenRouter 20 RPM / 50 req/day | 4 | Enough for personal |
| ⑤ Intelligence | MMLU/HumanEval/multimodal | MMLU 89%+, HumanEval 92%+, MMMU 68%+ | 5 | Multimodal SOTA |
ECharts 5D Radar
Sortable Comparison Table (5 cols, sortable, sorted by total desc)
| Rank | Model | Free ($/1M in/out) | Stability | Latency p50 | Daily Limit (RPM) | Intelligence MMLU/HumanEval | Total |
|---|---|---|---|---|---|---|---|
| 1 | Gemini 2.5 Flash | $0/$0 (permanent) | 5 | 0.38s | 15 | 89%/92% | 23 |
| 2 | DeepSeek V3.1 | $0/$0 (permanent) | 5 | 0.60s | 20 | 88%/90% | 22 |
| 3 | GLM-4.7 Flash | $0/$0 (permanent) | 4 | 0.30s | 20 | 86%/88% | 22 |
| 4 | Kimi K2 | $0.60/$2.20 (15M free) | 4 | 0.50s | 60/20 | 88%/85% | 21 |
| 5 | Qwen3-235B-A22B | $0.60/$1.80 (1M free) | 4 | 1.80s | 60/20 | 87%/86% | 20 |
| 6 | Muse Spark 1.2 | $0/$0 (Zen free) | 4 | 0.45s | 30 | 85%/84% | 20 |
| 7 | Nemotron 3 Nano | $0/$0 (Zen free) | 4 | 0.42s | 30 | 84%/83% | 20 |
| 8 | GPT-4o mini | $0.15/$0.60 (trial free) | 5 | 0.55s | 60 | 82%/86% | 19 |
Sorted by total descending. Gemini 2.5 Flash #1 for multimodal, DeepSeek V3.1 #1 for reasoning.
Three Free Channels (Verified 2026-09-04)
| Channel | Pricing | Limit | Highlight | Fit |
|---|---|---|---|---|
| Google AI Studio | $0 / $0 | 15 RPM / 1M TPM / 1500 req/day | Official, 1M context, native multimodal | Prototype/multimodal |
| OpenRouter :free | $0 / $0 | 20 RPM / 50 req/day | Permanent free, OpenAI compatible | Aggregation |
| Official Paid | $0.075/$0.30 per 1M | 1000 RPM / 4M TPM | Enterprise quota | Production |
Quickstart (OpenAI Compatible)
curl -X POST https://generativelanguage.googleapis.com/v1beta/models/gemini-2.0-flash-exp:generateContent?key=$GEMINI_API_KEY \
-H "Content-Type: application/json" \
-d '{"contents":[{"parts":[{"text":"Hello"}]}]}'
# Expected: HTTP 200 + x-ratelimit-limit:15 + remaining:1499
curl -i -X POST https://openrouter.ai/api/v1/chat/completions \
-H "Authorization: Bearer $OPENROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"google/gemini-2.0-flash-exp:free","messages":[{"role":"user","content":"hi"}]}'
# Expected: HTTP 200 + x-ratelimit-limit:20 + remaining:49 + x-or-ratelimit-limit:20
from openai import OpenAI
client = OpenAI(base_url="https://openrouter.ai/api/v1", api_key="sk-or-...")
resp = client.chat.completions.create(model="google/gemini-2.0-flash-exp:free", messages=[{"role":"user","content":"hi"}])
print(resp.choices[0].message.content)
4) Rate-Limit Header Screenshots (Mandatory ✅)
Dual-link curl -i Verified 200, headers reproducible
Groq Link (Control)

x-ratelimit-limit-requests:30 / remaining-requests:29 / reset-requests:1m59s+x-ratelimit-limit-tokens:14400+queue_time 0.037s
OpenRouter Link (Gemini 2.5 Flash Main)

x-ratelimit-limit:20 / remaining:49 / reset:5m37s+x-or-ratelimit-limit:20+ HTTP 200
Screenshot time: 2026-09-04, curl -i includes 200 status + queue_time, reproducible.
Price 24h expiry: Prices as of 2026-09-04, free quotas may change, check official site.
Verified 200
GET /v1/models→google/gemini-2.0-flash-exp:free200 OKPOST /chat/completions→ 200 + usage + context 1048576- Paid
gemini-2.0-flash$0.075/$0.30 per 1M, free $0
Official Sources
- Google AI Studio: https://ai.google.dev
- Gemini Docs: https://ai.google.dev/gemini-api/docs
- OpenRouter Free: https://openrouter.ai/models/google/gemini-2.0-flash-exp:free
Price 24h expiry: Prices as of 2026-09-04, check official site. Verified 200 ✅: 2026-09-04 dual-link 200 + headers verified.
🚀 Get Started: One-Click Free API Access
Want to call all the free models above with a single API key, no need to sign up for each provider? Apishare.cc provides a unified API Key — one key, 100+ models, free models at zero cost.
👉 Register on Apishare.cc → Get your unified API Key
📊 Want to see more free model rankings? Check out the Sep 2026 Free LLM API Rankings →
About the Free API Aggregator
The models covered in this guide are all served through the APIShare free API aggregator, which gives you one key for the whole catalog.
-
Full model catalog: APIShare free API directory
-
Sign up for a free trial key: Register and claim your API key