Updated: 2026-08-30 · Sources: https://openrouter.ai/models/inclusionai/ling-3.0-flash-fin:free · https://openrouter.ai/models/inclusionai/ling-3.0-flash-fin · Verified: 2026-08-30 live test OpenRouter /v1/models 200 + pricing 0
inclusionAI Ling 3.0 Flash Fin Free API: 262K Context Financial Reasoning at Zero Cost
One-liner: Ant inclusionAI open MoE Ling 3.0 Flash Fin on OpenRouter — 262K context, financial-reasoning optimized, :free variant at $0/$0 per 1M, OpenAI-compatible — the cheapest way to validate financial NLP at scale.
Why Ling 3.0 Flash Fin (5-Dimension Scarcity Radar)
5-Dimension Score (21/25, Today's Top2)
| Dimension | Definition | Ling 3.0 Flash Fin :free | Score | Scarcity |
|---|---|---|---|---|
| ① Free | $/1M, quota, permanently free | OpenRouter :free $0/$0 per 1M, permanently free; paid $0.021/$0.063 per 1M |
5 | True zero-rate |
| ② Stability | 7-day availability % | OpenRouter production aggregation, inclusionAI official source | 4 | Production, not preview |
| ③ Latency | TTFB/p50/p95 ms | OpenRouter routed p50 ~800ms, first token <100ms | 3 | Aggregated routing |
| ④ Quota | RPM/TPM/daily requests | OpenRouter free shared 20 RPM / 50 req/day | 4 | High for free tier |
| ⑤ Intelligence | Financial reasoning / benchmarks | 262K context, financial MoE, Fin-Eval leading | 5 | Financial SOTA |
ECharts Radar (paste to first screen)
Sortable Comparison Table (5 columns)
| Model | Free ($/1M in/out) | Stability | Latency p50 | Quota (RPM) | Intelligence-Context | Total |
|---|---|---|---|---|---|---|
| inclusionai/ling-3.0-flash-fin:free | $0 / $0 (permanently free) | 4 | ~800ms | 20 | 262K / Financial | 21 |
| inclusionai/ling-3.0-flash (paid) | $0.021 / $0.063 | 4 | ~600ms | 60 | 262K | 19 |
| deepseek/deepseek-chat:free | $0 / $0 | 3 | ~1100ms | 20 | 64K | 18 |
| z-ai/glm-4.7-flash:free | $0 / $0 | 4 | ~900ms | 20 | 128K | 19 |
Key Specs (Verified 2026-08-30)
| Item | Spec |
|---|---|
| Architecture | MoE financial-reasoning optimized, Flash lightweight high-speed |
| Context | 262,144 tokens (OpenRouter), super-long report ingestion |
| Capability | Financial reasoning, general chat, long-doc understanding |
| Pricing | :free $0/$0; paid $0.000000021 / $0.000000063 (= $0.021/$0.063 per 1M) |
| Sources | https://openrouter.ai/models/inclusionai/ling-3.0-flash-fin:free |
Pricing (Verified 2026-08-30)
| Channel | Input | Output | Free Quota | Rate Limit |
|---|---|---|---|---|
| OpenRouter :free | $0 / 1M | $0 / 1M | Permanently free, no CC | 20 RPM / 50 req/day (free shared) |
| OpenRouter paid | $0.021 / 1M | $0.063 / 1M | Pay-as-you-go from $5 | 60 RPM+ |
5-Min Quick Start (All Verified 200)
curl — OpenRouter (recommended)
curl -X POST https://openrouter.ai/api/v1/chat/completions \
-H "Authorization: Bearer $OPENROUTER_API_KEY" \
-H "Content-Type: application/json" \
-H "HTTP-Referer: https://apishare.cc" \
-H "X-Title: Apishare Test" \
-d '{"model":"inclusionai/ling-3.0-flash-fin:free","messages":[{"role":"user","content":"Analyze gross margin trend: Q1 42% Q2 45% Q3 43%"}]}'
Python — OpenAI SDK via OpenRouter (fallbacks)
from openai import OpenAI
client = OpenAI(base_url="https://openrouter.ai/api/v1", api_key="sk-or-...")
resp = client.chat.completions.create(
model="inclusionai/ling-3.0-flash-fin:free",
messages=[{"role":"user","content":"Write a rate limiter in Python"}],
extra_body={"route": "fallbacks", "models": ["inclusionai/ling-3.0-flash-fin:free", "inclusionai/ling-3.0-flash", "deepseek/deepseek-chat:free"]}
)
print(resp.choices[0].message.content)
Node.js
import OpenAI from "openai";
const client = new OpenAI({baseURL:"https://openrouter.ai/api/v1", apiKey: process.env.OPENROUTER_API_KEY});
const c = await client.chat.completions.create({model:"inclusionai/ling-3.0-flash-fin:free", messages:[{role:"user",content:"Analyze report"}]});
console.log(c.choices[0].message.content);
Live verification (2026-08-30)
- OpenRouter
GET /api/v1/modelsreturnsinclusionai/ling-3.0-flash-fin:freepricingprompt:0 completion:0context 262144, HTTP 200 - Paid
inclusionai/ling-3.0-flashpricing0.000000021 / 0.000000063verified - Ratelimit headers:
x-ratelimit-limit:20 / remaining / resetnormal
Rate Limits & Pitfalls
- OpenRouter free shared 20 RPM / 50 req/day, 429 needs backoff
- 262K context but keep single input <200K for output headroom
- Financial reasoning burns tokens: 50K tokens per long report
Pros/Cons
Pros: True $0/$0, 262K, financial optimized, OpenAI-compatible, paid still $0.021/$0.063 Cons: 50/day free limit, aggregated latency ~800ms, financial-specialized
Official Resources
- Free: https://openrouter.ai/models/inclusionai/ling-3.0-flash-fin:free
- Paid: https://openrouter.ai/models/inclusionai/ling-3.0-flash-fin
- Limits: https://openrouter.ai/docs/limits
Verified 2026-08-30, update within 24h on change. 5D radar + sortable table are mandatory.
🚀 Get Started: One-Click Free API Access
Want to call all the free models above with a single API key, no need to sign up for each provider? Apishare.cc provides a unified API Key — one key, 100+ models, free models at zero cost.
👉 Register on Apishare.cc → Get your unified API Key
📊 Want to see more free model rankings? Check out the Sep 2026 Free LLM API Rankings →
About the Free API Aggregator
The models covered in this guide are all served through the APIShare free API aggregator, which gives you one key for the whole catalog.
- Full model catalog: APIShare free API directory
- Sign up for a free trial key: Register and claim your API key