Sulat.com
AI models
Kilo Gateway logo

Model details

inclusionAI: Ling 3.0 Flash Fin

Ling 3.0 Flash Fin is a finance-tuned variant of the Ling 3.0 Flash reasoning model. It uses a hybrid-linear sparse mixture-of-experts architecture with 124 billion total parameters and activates 5.1 billion parameters for each token, preserving general reasoning ability while specializing it for financial work.

This model is designed for complex investment workflows, including multi-step research, financial document review, and long-horizon planning. Its large context supports extended documents and cross-source analysis, while tool calling, structured outputs, and built-in reasoning make it practical for financial-data agents and internal research systems. Public benchmark evidence remains limited: one independent tracker lists only three provisional sourced results, covers only agentic performance, and assigns no overall comparative rank.

Kilo Gatewayinclusionai/ling-3.0-flash-finling

Quick Info

Powered by
Provider
Kilo Gateway
Model key
inclusionai/ling-3.0-flash-fin
Release date
Aug 27, 2026
Last updated
Aug 27, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.06
Output token cost
$0.18

Limits

Output tokens
235,929 tokens
Context window
262,144 tokens

Transparent token rates

Compare inclusionAI: Ling 3.0 Flash Fin pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about inclusionAI: Ling 3.0 Flash Fin

Kilo Gateway

CoverageBenchmark

BenchLM's profile page treats Ling 3.0 Flash Fin as tracked but not publicly ranked, with only three provisional sourced benchmark rows available and no comparative rank assigned as of September 4, 2026. The spec sheet lists a 262K-token context window sourced from Vercel AI Gateway availability, marks the API model ID Coverage is split into eight categories (agentic, coding, reasoning, knowledge, math, multilingual, multimodal, instruction following), of which only agentic has any provisional data while the remaining seven are listed as not measured. The decision snapshot shows the capability field is unranked with a field median of

Kilo Gateway

Coverage

Vercel's changelog, as recapped by CreateWith, lists Ling 3.0 Flash Fin from Inclusion AI as available through Vercel AI Gateway and free to use through September 25, 2026. The model combines a 256K-token context window with reasoning, function calling, and up to 32K tokens of output, specifications that Vercel positio Developers integrating through Vercel must pick between two model identifiers, and the free window lets small teams prototype document-review assistants, internal research tools, and financial-data agents before committing to paid usage. The recap cautions that any finance model output should still be checked against o

Kilo Gateway

Coverage

Model Pulse tracks InclusionAI's Ling 3.0 Flash Fin across multiple providers and confirms its availability through Kilo Gateway with a listed price of $0.06 per million input tokens and $0.18 per million output tokens, plus a $0.012 cache-read rate. The Kilo Gateway variant exposes a 262K-token context window with rea According to the recent changes log, the Ling 3.0 Flash Fin listing was first surfaced via Kilo on September 3, 2026, with an earlier provider entry dated August 27, 2026, and a Vercel follow-up on August 28, 2026. Best published input and output price history shows the lowest listed price holding flat at $0.06/$0.18 p

Videos about inclusionAI: Ling 3.0 Flash Fin

More models around inclusionAI: Ling 3.0 Flash Fin