DeepSeek-R1-Distill-7B
by DeepSeek
· DeepSeek R1 family · released Jan 20, 2025
Open weights
Active
Qwen-distilled R1 reasoning at 7B.
This model has no first-party API. Open-weight release: self-host from published weights.
Specifications
| Model type | Text LLM |
|---|---|
| Architecture | Dense |
| Parameters | 7 |
| Context window | 131.1K tokens |
| Max output | 8.2K tokens |
| Knowledge cutoff | Aug 1, 2024 |
| License | MIT License |
| Modalities | Text input Text output |
| Tags | Distilled Reasoning Small LM |
| End of life | Not disclosed |
| Last verified | Aug 20, 2026 |