Hunyuan Large
Hunyuan Large is Tencent’s open-weight 389B-parameter mixture-of-experts language model with 52B active parameters. The release publishes pretrained, instruction-tuned, and FP8 instruction checkpoints together with training and inference code. Grouped-query and cross-layer attention reduce key-value cache requirements for long documents. The pretrained model supports 256K context, while the instruction model uses 128K; this page records the conservative conversational limit. Its custom Tencent license governs downloadable weights, and local or Tencent Cloud TI deployments support customization. The November 18 instruction refresh is grouped as a variant.
2024-11-05
389B total, 52B active
Mixture-of-Experts Transformer with grouped-query and cross-layer attention
Tencent Hunyuan Community License
Specifications
- Parameters
- 389B total, 52B active
- Architecture
- Mixture-of-Experts Transformer with grouped-query and cross-layer attention
- License
- Tencent Hunyuan Community License
- Context Window
- 131,072 tokens
- Type
- text
- Modalities
- text
Benchmark Scores
Advanced Specifications
- Model Family
- Hunyuan
- API Access
- Not Available
- Chat Interface
- Not Available
- Multilingual Support
- Yes
- Variants
- Hunyuan-A52B-PretrainHunyuan-A52B-InstructHunyuan-A52B-Instruct-FP82024-11-18 Instruct refresh
Capabilities & Limitations
- Capabilities
- text generationinstruction followingChinese and Englishlong-context processingreasoning
- Known Limitations
- Custom Tencent license appliesFull weight memory requirements are substantial256K applies to pretrained checkpoint; Instruct is 128K
- Notable Use Cases
- private multilingual assistantslong-document researchmodel fine-tuning