InclusionAI Releases Ling 3.0 Flash Fin, OpenRouter Adds a Route

InclusionAI's official model card dates Ling 3.0 Flash Fin to 3 September 2026. OpenRouter then listed paid and zero-price routes for the 124-billion-parameter finance model.
Synthesized via Fish Audio S2.1 Pro in British English. Natural editorial summary, not verbatim reading.
Key Takeaways
- check_circleInclusionAI published the official Ling 3.0 Flash Fin model card and BF16 weights on 3 September 2026.
- check_circleThe model has 124 billion total parameters, about 5.1 billion active parameters and a 256K context window.
- check_circleAIMI first saw the standard OpenRouter route at 16:30:02 SAST, 8 hours, 20 minutes and 9 seconds after the official model-card timestamp.
- check_circleOpenRouter also lists a zero-price :free route. Free access and rate limits can change independently of the model release.
The official model release
InclusionAI published the Ling-3.0-flash-Fin model card on Hugging Face at 06:09:53 UTC, or 08:09:53 SAST, on 3 September 2026. The card describes it as the first finance-enhanced model in the Ant Ling family, developed by Ant Group with financial institutions and domain experts.
The same model card lists the current BF16 checkpoint and links to the model files. That is direct evidence of a maker release and weights publication, not just a provider catalogue entry.
What Ling 3.0 Flash Fin is built for
Ling 3.0 Flash Fin is a mixture-of-experts model with 124 billion total parameters and about 5.1 billion activated per token. Its 256K context window is intended for long financial documents and workflows that combine retrieval, evidence review, calculations, modelling and report preparation.
The model card names financial research, source-grounded search, multi-document reasoning, valuation and spreadsheet workflows. It reports evaluation across FinFIRST, FinSearchComp Verified, FinCRAFT, Finance Agent, APEX-Agents, SpreadsheetBench and τ³-Banking, but does not provide enough detail on this page to treat those results as an independent benchmark audit.
OpenRouter picked it up later
AIMI first saw the standard OpenRouter route, inclusionai/ling-3.0-flash-fin, at 16:30:02 SAST. That was exactly 8 hours, 20 minutes and 9 seconds after the official model-card timestamp. OpenRouter also exposes a zero-price inclusionai/ling-3.0-flash-fin:free route, first observed by AIMI on 27 August at 18:16:31 SAST.
The standard route lists a 262,144-token context window, 235,929-token maximum output, reasoning, tools and function calling. The :free route has a 32,768-token maximum output and zero prompt and completion pricing in the current catalogue. These are provider route details, not extra maker releases.
Use the finance claims with care
The model card says complex financial workflows still need further validation and that investment conclusions require professional review. It does not constitute investment advice. A large context window and tool support can help with research, but they do not make generated analysis reliable by themselves.
AZ Labs does not sell an InclusionAI subscription. The relevant product is the AZ Labs AI Gateway, which gives applications one integration point for model routing and provider changes. Test the exact route, prompts, tools and rate limits before using it for financial work or other high-impact decisions.
Frequently Asked Questions
What was released?
InclusionAI released Ling 3.0 Flash Fin, a finance-enhanced model in the Ant Ling family, with a BF16 checkpoint and model files on Hugging Face.
Is Ling 3.0 Flash Fin free?
OpenRouter currently lists a zero-price :free route, but free endpoints are rate limited and pricing or availability can change.
What context and output limits are listed?
The route lists a 262,144-token context window and a maximum output of 32,768 tokens.
Was this an official InclusionAI release?
Yes. InclusionAI's official Hugging Face model card was created on 3 September 2026 and lists the BF16 checkpoint and model files.
What are the exact OpenRouter routes?
The standard route is inclusionai/ling-3.0-flash-fin. OpenRouter also lists inclusionai/ling-3.0-flash-fin:free, with zero prompt and completion pricing when checked.
Does the route support tools?
OpenRouter metadata lists reasoning, tools and function calling for the standard route. Test the exact tool schema and workflow before relying on it in production.
Related Frontier Models & Releases
Explore verified specifications, benchmark results, and route pricing across alternative models in this class.
Ling 3.0 Flash Runs in Grok CLI Through OpenRouter
A practical model-routing experiment puts InclusionAI's Ling 3.0 Flash inside Grok CLI, with the request verified through OpenRouter and the same route working in Pi.
OpenRouter Lists InclusionAI's Ling 3.0 Tiny as a Free Route
OpenRouter now lists InclusionAI's Ling 3.0 Tiny as a zero-price route with a 262K-token context window, switchable thinking and instant modes, and a 32K maximum output.
DeepSeek releases V4.1 Flash with native multimodal vision, 552B MoE and lower API rates
DeepSeek has launched DeepSeek-V4.1-Flash, a 552B-parameter mixture-of-experts model featuring a novel Causal Encoder–Decoder architecture with just 8B active input and 16B active output parameters. The release brings native vision, compresses KV cache storage by up to 8x, and slashes off-peak API pricing to $0.15 per million input tokens.