inclusionAI: Ling 3.0 Flash
inclusionai/ling-3.0-flash
Ling 3.0 Flash (inclusionAI) is a low-cost open-weight model priced near the bottom of the board, at $0.021/$0.063 per million input/output tokens, served by 2 independent hosts. For the money it delivers #96 of 241 on the TryAii Score — among the best score-per-dollar in its class. Its best showing is #23 on AA-Omniscience-Hallucination-Rate. It's among the fastest models we track, at roughly 58-224 tokens/sec.
Prompt $/1M
$0.021
Completion $/1M
$0.063
Context
262,144 tok
Best provider
354 tok/s · Artificial Analysis
Specs
- Model id
- inclusionai/ling-3.0-flash
- Display name
- inclusionAI: Ling 3.0 Flash
- Modality
- text->text
- Tokenizer
- Other
- Instruct type
- —
- Context length
- 262,144
- Max completion
- 32,768
- Prompt $/1M
- $0.021
- Completion $/1M
- $0.063
- Free tier
- no
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Artificial Analysis | 354 |
| Novitafastest | 95 |
| DeepInfra | 22 |
| inclusionaifastest | — |