Ling-2.5-1T
Ling-2.5-1T updates Ant Group’s trillion-parameter language family with hybrid linear attention for efficient long contexts. It combines Multi-head Latent Attention and Lightning Linear Attention in a 1:7 ratio with sparse experts. This instant-model tier targets concise reasoning, preference alignment, creative writing, and native agent interaction, with a native 256K context extendable to one million tokens. Both downloadable MIT weights and historical hosted access are documented. The first-party 2.5 APIs were retired September 30; this page preserves the released model identity.
2026-02-16
1T total, 63B active
Hybrid MLA/Lightning Linear Attention Mixture of Experts
MIT
Specifications
- Parameters
- 1T total, 63B active
- Architecture
- Hybrid MLA/Lightning Linear Attention Mixture of Experts
- License
- MIT
- Context Window
- 262,144 tokens
- Type
- text
- Modalities
- text
Benchmark Scores
Advanced Specifications
- Model Family
- Ling
- API Access
- Not Available
- Chat Interface
- Not Available
Capabilities & Limitations
- Capabilities
- reasoningcodingtool useinstruction following
- Known Limitations
- Inactive experts still require substantial weight storage.Tool use requires an external execution environment and appropriate chat template.
- Notable Use Cases
- coding agentsenterprise assistantslong-document analysis