Ling 3.1 Flash
inclusionAI · released 2 Oct 2026
InputFreeper 1M tokens
OutputFreeper 1M tokens
Context262Ktokens
WeightsClosed
context on a 1K–2M scale
About
Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total.
Details
- Lab
- inclusionAIopenrouter.ai · 3 Oct 2026
- Context
- 262Kopenrouter.ai · 3 Oct 2026
- Input price
- Freeopenrouter.ai · 3 Oct 2026
- Output price
- Freeopenrouter.ai · 3 Oct 2026
- Max output
- 32,768 tokensopenrouter.ai · 3 Oct 2026
- Inputs
- textopenrouter.ai · 3 Oct 2026
- Open weights
- Noopenrouter.ai · 3 Oct 2026
More longest-context models
See the listListed on Inferse
- Longest-Context Models in 2026466 listed
- Newest Models in 2026466 listed
- Models with Tool Calling in 2026398 listed
- Reasoning Models in 2026332 listed
- Free Models in 202622 listed
- inclusionAI Models: Pricing per 1M Tokens and Context (2026)5 listed
Sources
- openrouter.ai/inclusionai/ling-3.1-flash· checked 3 Oct 2026


