Free LLM API Leaderboard · 2026 Live Test

Daily probes at 06:00 UTC; deep quota re-check at 03:00 UTC. Composite score = 50% availability + 30% speed + 20% rate-limit headroom.

📊 Current Rankings

# Provider Availability Latency (ms) RPM Daily Limit Models Composite Score
1
nvidia-nim
● Online 280 60 0 83 99.2 详情 ›
2
agnes-free
● Online 320 60 0 10 99.0 详情 ›
3
groq
● Online 85 30 0 13 89.7 详情 ›
4
cloudflare-free
● Online 250 30 0 2 89.2 详情 ›
5
mistral
● Online 180 20 0 39 86.1 详情 ›
6
cohere-trial
● Online 210 20 0 12 86.0 详情 ›
7
gemini
● Online 120 15 0 10 84.6 详情 ›
8
opencode-zen
● 探针排查中 639 0 0 8 0.0 详情 ›

📝 Daily Changelog

2026-08-22
【更正】Gemini 渠道模型计数统一为实测入库口径(28→10):此前榜单页展示的 Gemini 模型数(主榜16+专项12=28)系 key 验证期候选分析口径,「模型数30」则来自上游 /v1beta/models 列表重复计数,均未实际写入渠道池。经生产库直读 + 全模型逐个冒烟复核,Gemini 渠道池实际入库 10 个对话文本模型(8 个 gemini text + 2 个 gemma),10/10 实测 HTTP 200。全站文章已同步更正:模型表改为 10 模型权威清单,专项区块移除,演示型号切换为 gemini-3.6-flash;gemini-2.5-* 对新用户返回 404 已下线并标注。
[Correction] Gemini channel model count unified to verified provisioned list (28→10): the previously displayed Gemini model count on the leaderboard (16 main + 12 specialized = 28) was a candidate-analysis snapshot from the key-validation phase, while the "30 models" figure came from duplicate counting of the upstream /v1beta/models listing; neither was ever provisioned into the channel pool. Verified via production DB direct read plus per-model smoke tests, the Gemini pool actually contains 10 conversational text models (8 gemini text + 2 gemma), all returning HTTP 200. All site articles have been updated accordingly: the model table now lists the authoritative 10-model roster, the specialized section is removed, the demo model is switched to gemini-3.6-flash, and gemini-2.5-* (404 for new users) has been delisted with an explicit note.
2026-08-22
Gemini 与 Mistral 两渠道正式入池上线:Gemini 10 个对话/文本模型实测入库并进入主榜综合评分(旗舰 gemini-3.7-flash、免费层最稳的 gemini-3.6-flash、gemma-4 开源双子等,Flash 系列最高 1M 输入 / 64K 输出;图像/TTS/向量等专项能力免费层配额为 0,未入池);Mistral 39 个 chat 类模型全量入池(mistral-large/medium/small、magistral、devstral、ministral、codestral 等),9 个 voxtral-* 音频模型标 audio 进专项区。详见 《Google Gemini 免费层》《Mistral 免费 API》
Gemini and Mistral officially joined the channel pool: 10 Gemini chat/text models verified into the main ranking (flagship gemini-3.7-flash, free-tier-stable gemini-3.6-flash, the open-source Gemma pair, Flash series up to 1M-token input / 64K output; specialist capabilities like image/TTS/embedding carry zero free-tier quota and are not pooled); Mistral ships 39 chat-class models into the pool (mistral-large/medium/small, magistral, devstral, ministral, codestral, etc.) plus 9 voxtral-* audio models tagged audio in the specialist block. See the Google Gemini Free Tier guide and the Mistral Free API guide.
2026-08-22
Groq 渠道池 13 个模型逐一实测通过并全量上线:主榜新增 openai/gpt-oss-120b、openai/gpt-oss-20b、qwen/qwen3.6-27b、groq/compound、groq/compound-mini、allam-2-7b;专项区新增 whisper-large-v3(-turbo)、orpheus TTS×2(待 Groq 条款激活)、gpt-oss-safeguard-20b、llama-prompt-guard×2;旧 llama-3.x 已移出渠道池。详见《Groq 免费极速推理 API》
Groq channel pool fully verified: 13 models tested live. Main ranking adds openai/gpt-oss-120b, openai/gpt-oss-20b, qwen/qwen3.6-27b, groq/compound, groq/compound-mini, allam-2-7b; specialist block adds whisper-large-v3(-turbo), Orpheus TTS ×2 (pending Groq ToS activation), gpt-oss-safeguard-20b, llama-prompt-guard ×2; legacy llama-3.x removed from the pool. See the Groq guide for details.

Every channel change is recorded here after live verification.

📈 Latency Trend (ms) — lower is better

📊 Availability Trend — higher is better

Data source: /api/free-llm-rankings/history (JSON, consumed by the chart).