inclusionai

inclusionAI: Ling 3.0 Flash Fin

inclusionai/ling-3.0-flash-fin

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

Prices updated: 10/7/2026

Input
4.6
credits / 1M tokens
Output
14
credits / 1M tokens
Context
262k
tokens
Messages per $1 USD
≈ 8,800
typical message: 1k tokens in + 500 out

Prices in credits per million tokens. 1 credit = $0.01 USD. Final price, billed in MXN at the exchange rate on payment day + VAT (CFDI). Subject to change without notice.

Additional prices

Cache (read): 0.9 /1M

Spec sheet

Inputs
text
Outputs
text
Max output tokens
32,768
Knowledge cutoff
—
Tokenizer
Other
Capabilities
Text / LLM
Tools
JSON
Reasoning

Use it in your code

One OpenAI-compatible endpoint. Switch models by changing one line.

curl https://api.deployalo.com/ai/gateway/v1/chat/completions \
  -H "Authorization: Bearer $DEPLOYALO_AI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"inclusionai/ling-3.0-flash-fin","messages":[{"role":"user","content":"Hola"}]}'

Ready to use it?

Top up a credit pack, create your API key with a spend cap and start in minutes.