Currently listed through these providers:
Model details
Ling 3.0 Flash Fin
Ling 3.0 Flash Fin is a finance-oriented large language model from InclusionAI that builds directly on the Ling 3.0 Flash base, adapting the same mixture-of-experts backbone for investment, research, and analytical workloads. The architecture follows Ling 3.0 Flash's MoE design, activating roughly 5.1 billion parameters per token out of a 124 billion parameter total, which keeps per-request compute low while preserving the breadth of a much larger expert pool. This finance-specific finetune reflects a broader trend of taking general MoE foundations and specializing them for vertical domains where domain vocabulary, numerical reasoning, and document structure matter more than open-ended chat fluency.
In practice, Ling 3.0 Flash Fin is positioned for long-context analysis of filings, reports, and market commentary, paired with tool-calling so the model can pull live data, run calculations, or chain into downstream agents. Reasoning support and tool use are both highlighted as core capabilities, making it a fit for workflows that blend narrative summarization with structured lookups rather than casual conversation. Teams evaluating it for financial assistants, research copilots, or agent pipelines should weigh the active-parameter efficiency against the larger total parameter count when planning latency, throughput, and serving costs.
Quick Info
Powered by- Provider
- Vercel AI Gateway
- Model key
- inclusionai/ling-3.0-flash-fin
- Release date
- Aug 27, 2026
- Last updated
- Aug 27, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
A provider subscription or plan supersedes token-based pricing for this model.
Limits
- Output tokens
- 32,000 tokens
- Context window
- 256,000 tokens