Ling 3.0 Flash Fast
Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers to complete more useful work within constrained token, latency, and serving-cost budgets.
text → text · made by InclusionAI
$0.06→$0.18per Mtok in / out
Novita AI
Sold by 1 way
| Seller | Lane | Rate |
|---|---|---|
| Novita AI aggregator | standard | $0.06per Mtok in$0.18per Mtok out$0.012per Mtok cachedapi.novita.ai · read 2026-08-25 |
About
Ling 3.0 Flash Fast — a text model from InclusionAI, sold by one company from $0.06 in and $0.18 out per million tokens.
It takes text and returns text, with a context window of 262,144 tokens. The catalogue files it under chat. Only Novita AI sells it, at $0.06 in and $0.18 out per million tokens.
Every current figure
- Maker
- InclusionAI
- Register
- model
- Takes
- text
- Returns
- text
- Context
- 262,144 tokens
- Sellers
- 1
- Maker's own price
- not read
- Price
- $0.06 in and $0.18 out per million tokens — Novita AI
Known as 1 name
inclusionai/ling-3.0-flash-fast