Free LLM API Leaderboard · 2026 Live Test

Daily probes at 06:00 UTC; deep quota re-check at 03:00 UTC. Composite score = 35% availability + 15% speed + 10% headroom + 40% model capability (large-model first).

📊 Current Rankings

# Provider Latency (ms) RPM Models Composite Score
1
nvidia-nim
1194 — 82 93.2 详情 ›
2
openrouter
621 — 19 92.1 详情 ›
3
groq
534 1000 10 81.2 详情 ›
4
agnes-free
5879 — 12 80.2 详情 ›
5
cohere-trial
451 20 13 69.5 详情 ›
6
cloudflare-free
1534 — 2 68.7 详情 ›
观察中 — 可用性 0(当前不可用,不建议生产使用)
0
mistral
285 — 35 40.4 详情 ›
0
gemini
564 — 14 38.4 详情 ›

📝 Daily Changelog

2026-08-22
【更正】Gemini 渠道模型计数统一为实测入库口径(28→10):此前榜单页展示的 Gemini 模型数(主榜16+专项12=28)系 key 验证期候选分析口径,「模型数30」则来自上游 /v1beta/models 列表重复计数,均未实际写入渠道池。经生产库直读 + 全模型逐个冒烟复核,Gemini 渠道池实际入库 10 个对话文本模型(8 个 gemini text + 2 个 gemma),10/10 实测 HTTP 200。全站文章已同步更正:模型表改为 10 模型权威清单,专项区块移除,演示型号切换为 gemini-3.6-flash;gemini-2.5-* 对新用户返回 404 已下线并标注。
[Correction] Gemini channel model count unified to verified provisioned list (28→10): the previously displayed Gemini model count on the leaderboard (16 main + 12 specialized = 28) was a candidate-analysis snapshot from the key-validation phase, while the "30 models" figure came from duplicate counting of the upstream /v1beta/models listing; neither was ever provisioned into the channel pool. Verified via production DB direct read plus per-model smoke tests, the Gemini pool actually contains 10 conversational text models (8 gemini text + 2 gemma), all returning HTTP 200. All site articles have been updated accordingly: the model table now lists the authoritative 10-model roster, the specialized section is removed, the demo model is switched to gemini-3.6-flash, and gemini-2.5-* (404 for new users) has been delisted with an explicit note.
2026-08-22
Gemini 与 Mistral 两渠道正式入池上线:Gemini 10 个对话/文本模型实测入库并进入主榜综合评分(旗舰 gemini-3.7-flash、免费层最稳的 gemini-3.6-flash、gemma-4 开源双子等,Flash 系列最高 1M 输入 / 64K 输出;图像/TTS/向量等专项能力免费层配额为 0,未入池);Mistral 39 个 chat 类模型全量入池(mistral-large/medium/small、magistral、devstral、ministral、codestral 等),9 个 voxtral-* 音频模型标 audio 进专项区。详见 《Google Gemini 免费层》 与 《Mistral 免费 API》。
Gemini and Mistral officially joined the channel pool: 10 Gemini chat/text models verified into the main ranking (flagship gemini-3.7-flash, free-tier-stable gemini-3.6-flash, the open-source Gemma pair, Flash series up to 1M-token input / 64K output; specialist capabilities like image/TTS/embedding carry zero free-tier quota and are not pooled); Mistral ships 39 chat-class models into the pool (mistral-large/medium/small, magistral, devstral, ministral, codestral, etc.) plus 9 voxtral-* audio models tagged audio in the specialist block. See the Google Gemini Free Tier guide and the Mistral Free API guide.
2026-08-22
Groq 渠道池 13 个模型逐一实测通过并全量上线:主榜新增 openai/gpt-oss-120b、openai/gpt-oss-20b、qwen/qwen3.6-27b、groq/compound、groq/compound-mini、allam-2-7b;专项区新增 whisper-large-v3(-turbo)、orpheus TTS×2(待 Groq 条款激活)、gpt-oss-safeguard-20b、llama-prompt-guard×2;旧 llama-3.x 已移出渠道池。详见《Groq 免费极速推理 API》。
Groq channel pool fully verified: 13 models tested live. Main ranking adds openai/gpt-oss-120b, openai/gpt-oss-20b, qwen/qwen3.6-27b, groq/compound, groq/compound-mini, allam-2-7b; specialist block adds whisper-large-v3(-turbo), Orpheus TTS ×2 (pending Groq ToS activation), gpt-oss-safeguard-20b, llama-prompt-guard ×2; legacy llama-3.x removed from the pool. See the Groq guide for details.

Every channel change is recorded here after live verification.

📈 Latency Trend (ms) — lower is better

Data source: /api/free-llm-rankings/history (JSON, consumed by the chart).