AZ Labs
Industry News3 September 20265 min read

InclusionAI Releases Ling 3.0 Flash Fin, OpenRouter Adds a Route

Model artwork for InclusionAI Ling 3.0 Flash Fin on an OpenRouter route
Inspect
OpenRouter Ling 3.0 Flash Fin route OpenRouter Ling 3.0 Flash Fin route© OpenRouter, used for news reporting

InclusionAI's official model card dates Ling 3.0 Flash Fin to 3 September 2026. OpenRouter then listed paid and zero-price routes for the 124-billion-parameter finance model.

smart_toyInclusionAIinclusionai/ling-3.0-flash-fin124B MoE total (5.1B active active)
verifiedFirst observed by AIMI: 2026-09-03 14:12:00 SAST
Context Windowarticle
262K tokens (262,144)
Max output: 33K tokens (32,768)
INInput Modalitiesinput
text input
descriptiontextattach_filefile
OUTOutput Contractoutput
text output
chattext
Route Pricingpayments
boltFree Route (Zero Cost)
Verified Model Capabilities & Tools
psychologyReasoning / ThinkingconstructionFunction Calling & Toolsdata_objectStructured Outputs (JSON)streamToken Streaming
Available Gateways:openrouter/inclusionai/ling-3.0-flash-fin
Listen to Article
Full Story
0:00 / 0:00

Synthesized via Fish Audio S2.1 Pro in British English. Natural editorial summary, not verbatim reading.

verified

Key Takeaways

  • check_circleInclusionAI published the official Ling 3.0 Flash Fin model card and BF16 weights on 3 September 2026.
  • check_circleThe model has 124 billion total parameters, about 5.1 billion active parameters and a 256K context window.
  • check_circleAIMI first saw the standard OpenRouter route at 16:30:02 SAST, 8 hours, 20 minutes and 9 seconds after the official model-card timestamp.
  • check_circleOpenRouter also lists a zero-price :free route. Free access and rate limits can change independently of the model release.

The official model release

InclusionAI published the Ling-3.0-flash-Fin model card on Hugging Face at 06:09:53 UTC, or 08:09:53 SAST, on 3 September 2026. The card describes it as the first finance-enhanced model in the Ant Ling family, developed by Ant Group with financial institutions and domain experts.

The same model card lists the current BF16 checkpoint and links to the model files. That is direct evidence of a maker release and weights publication, not just a provider catalogue entry.

What Ling 3.0 Flash Fin is built for

Ling 3.0 Flash Fin is a mixture-of-experts model with 124 billion total parameters and about 5.1 billion activated per token. Its 256K context window is intended for long financial documents and workflows that combine retrieval, evidence review, calculations, modelling and report preparation.

The model card names financial research, source-grounded search, multi-document reasoning, valuation and spreadsheet workflows. It reports evaluation across FinFIRST, FinSearchComp Verified, FinCRAFT, Finance Agent, APEX-Agents, SpreadsheetBench and τ³-Banking, but does not provide enough detail on this page to treat those results as an independent benchmark audit.

OpenRouter picked it up later

AIMI first saw the standard OpenRouter route, inclusionai/ling-3.0-flash-fin, at 16:30:02 SAST. That was exactly 8 hours, 20 minutes and 9 seconds after the official model-card timestamp. OpenRouter also exposes a zero-price inclusionai/ling-3.0-flash-fin:free route, first observed by AIMI on 27 August at 18:16:31 SAST.

The standard route lists a 262,144-token context window, 235,929-token maximum output, reasoning, tools and function calling. The :free route has a 32,768-token maximum output and zero prompt and completion pricing in the current catalogue. These are provider route details, not extra maker releases.

Use the finance claims with care

The model card says complex financial workflows still need further validation and that investment conclusions require professional review. It does not constitute investment advice. A large context window and tool support can help with research, but they do not make generated analysis reliable by themselves.

AZ Labs does not sell an InclusionAI subscription. The relevant product is the AZ Labs AI Gateway, which gives applications one integration point for model routing and provider changes. Test the exact route, prompts, tools and rate limits before using it for financial work or other high-impact decisions.

Frequently Asked Questions

What was released?

InclusionAI released Ling 3.0 Flash Fin, a finance-enhanced model in the Ant Ling family, with a BF16 checkpoint and model files on Hugging Face.

Is Ling 3.0 Flash Fin free?

OpenRouter currently lists a zero-price :free route, but free endpoints are rate limited and pricing or availability can change.

What context and output limits are listed?

The route lists a 262,144-token context window and a maximum output of 32,768 tokens.

Was this an official InclusionAI release?

Yes. InclusionAI's official Hugging Face model card was created on 3 September 2026 and lists the BF16 checkpoint and model files.

What are the exact OpenRouter routes?

The standard route is inclusionai/ling-3.0-flash-fin. OpenRouter also lists inclusionai/ling-3.0-flash-fin:free, with zero prompt and completion pricing when checked.

Does the route support tools?

OpenRouter metadata lists reasoning, tools and function calling for the standard route. Test the exact tool schema and workflow before relying on it in production.

Explore verified specifications, benchmark results, and route pricing across alternative models in this class.

Primary Sources

Share this articlePost on X
arrow_backBack to all news