Ling-3.0-flash (free)
inclusionai/ling-3.0-flash:free
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Prompt $/1M
free
Completion $/1M
free
Context
262,144 tok
Best provider
59 tok/s · Novita
Specs
- Model id
- inclusionai/ling-3.0-flash:free
- Display name
- Ling-3.0-flash (free)
- Modality
- text->text
- Tokenizer
- Other
- Instruct type
- —
- Context length
- 262,144
- Max completion
- 32,768
- Prompt $/1M
- free
- Completion $/1M
- free
- Free tier
- yes
Capabilities
- Function callingno
- JSON modeno
- System promptno
- Visionno
- Moderatedno
Benchmark scores
Each axis is one benchmark, normalized to the best model in the field.
No benchmark data for this model.
Speed by provider
| Provider | tok/s |
|---|---|
| Novitafastest | 59 |