Allen Institute for AI logo

DR Tulu

Allen Institute for AIOpen WeightsPending Human Review

DR Tulu is Ai2’s open deep-research language model and end-to-end training recipe for producing long-form grounded reports. The released 8B checkpoint builds on Qwen3-8B with supervised training and Reinforcement Learning with Evolving Rubrics. It learns to search, gather evidence, and synthesize reports through an external research harness. Ai2 releases checkpoints under the RL ReSearch organization, along with training data, tool stack, and evaluation artifacts. Report-quality scores measure the complete model-plus-search system and are not general standalone factual accuracy.

2025-11-18
8B
Qwen3-derived decoder-only Transformer
Apache-2.0

Specifications

Parameters
8B
Architecture
Qwen3-derived decoder-only Transformer
License
Apache-2.0
Type
text
Modalities
text

Benchmark Scores

Advanced Specifications

Model Family
Tülu
Finetuned From
Qwen3-8B
API Access
Not Available
Chat Interface
Not Available
Variants
SFTRLNo-RLER ablation

Capabilities & Limitations

Capabilities
deep researchtool uselong-form synthesis
Known Limitations
Requires external search and research toolsReports and citations may be inaccurate
Notable Use Cases
literature researchgrounded report draftsresearch-agent training
Tool Use Support
Yes

Related Models