Ring-2.5-1T
Ring-2.5-1T updates Ant Group’s trillion-parameter language family with hybrid linear attention for efficient long contexts. It combines Multi-head Latent Attention and Lightning Linear Attention in a 1:7 ratio with sparse experts. This thinking tier strengthens deep reasoning and sustained agent execution through fully asynchronous agentic reinforcement learning, with a native 128K context extendable to 256K. Both downloadable MIT weights and historical hosted access are documented. The first-party 2.5 APIs were retired September 30; this page preserves the released model identity.
2026-02-16
1T total, 63B active
Hybrid MLA/Lightning Linear Attention Mixture of Experts
MIT
Specifications
- Parameters
- 1T total, 63B active
- Architecture
- Hybrid MLA/Lightning Linear Attention Mixture of Experts
- License
- MIT
- Context Window
- 131,072 tokens
- Type
- text
- Modalities
- text
Benchmark Scores
Advanced Specifications
- Model Family
- Ring
- API Access
- Not Available
- Chat Interface
- Not Available
Capabilities & Limitations
- Capabilities
- reasoningcodingtool useinstruction following
- Known Limitations
- Inactive experts still require substantial weight storage.Tool use requires an external execution environment and appropriate chat template.
- Notable Use Cases
- coding agentsenterprise assistantslong-document analysis