DR Tulu
DR Tulu is Ai2’s open deep-research language model and end-to-end training recipe for producing long-form grounded reports. The released 8B checkpoint builds on Qwen3-8B with supervised training and Reinforcement Learning with Evolving Rubrics. It learns to search, gather evidence, and synthesize reports through an external research harness. Ai2 releases checkpoints under the RL ReSearch organization, along with training data, tool stack, and evaluation artifacts. Report-quality scores measure the complete model-plus-search system and are not general standalone factual accuracy.
2025-11-18
8B
Qwen3-derived decoder-only Transformer
Apache-2.0
Specifications
- Parameters
- 8B
- Architecture
- Qwen3-derived decoder-only Transformer
- License
- Apache-2.0
- Type
- text
- Modalities
- text
Benchmark Scores
Advanced Specifications
- Model Family
- Tülu
- Finetuned From
- Qwen3-8B
- API Access
- Not Available
- Chat Interface
- Not Available
- Variants
- SFTRLNo-RLER ablation
Capabilities & Limitations
- Capabilities
- deep researchtool uselong-form synthesis
- Known Limitations
- Requires external search and research toolsReports and citations may be inaccurate
- Notable Use Cases
- literature researchgrounded report draftsresearch-agent training
- Tool Use Support
- Yes