Ling 3.1 Flash
Ling 3.1 Flash is InclusionAI’s hosted hybrid reasoning language model for coding, multi-step analysis, and agents that call tools. Its mixture-of-experts network has 560B total parameters and activates 25B per token. The verified September 30 availability announcement provides a 262K context through Vercel AI Gateway, suitable for long documents, repository analysis, and extended task histories. Standard and temporary free API identifiers represent the same model. Public weights were not verified at the audit date, so hosted availability is recorded without assuming a downloadable license or checkpoint.
2026-09-30
560B total, 25B active
Mixture of Experts
Proprietary
Specifications
- Parameters
- 560B total, 25B active
- Architecture
- Mixture of Experts
- License
- Proprietary
- Context Window
- 262,144 tokens
- Type
- text
- Modalities
- text
Benchmark Scores
Advanced Specifications
- Model Family
- Ling
- API Access
- Available
- Chat Interface
- Not Available
- Variants
- inclusionai/ling-3.1-flashinclusionai/ling-3.1-flash-free
Capabilities & Limitations
- Capabilities
- hybrid reasoningcodingtool uselong context
- Known Limitations
- Public weights were not verified at the audit dateThe documented gateway context can differ from other deploymentsAgent actions require an external tool harness and validation
- Notable Use Cases
- coding agentsmulti-step document analysislong-running tool workflows
- Tool Use Support
- Yes