deepseek-ai/DeepSeek-R1-Distill-Llama-8B
Parameters8B
Context128K
Input → outputtext → text
Licensemit
PrecisionBF16
ArchitectureLlamaForCausalLM
ReleasedJan 20, 2025
Downloads164.2K
Other sizes and formats in the family
| Model | Params | Precision | Downloads |
|---|---|---|---|
| deepseek-ai/DeepSeek-R1-Zerofp8 | 685B | F8_E4M3 | 8K |
| deepseek-ai/DeepSeek-R1fp8 | 684B | F8_E4M3 | 1.1M |
| deepseek-ai/DeepSeek-R1-Distill-Llama-70B | 70.6B | BF16 | 56.2K |
| deepseek-ai/DeepSeek-R1-Distill-Qwen-32B | 32.8B | BF16 | 409K |
| deepseek-ai/DeepSeek-R1-Distill-Qwen-14B | 14.8B | BF16 | 328.5K |
| deepseek-ai/DeepSeek-R1-Distill-Qwen-7B | 7.6B | BF16 | 322.6K |
| deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B | 1.8B | BF16 | 1.2M |
This release in the news
No coverage of this exact release yet.
Recent news about DeepSeek models
- Oct 11, 2026 · Opinion / analysisPSA: DeepSeek V4.1 Flash habitually exfiltrates API keys. It is dangerously misaligned and may be hazardous to use
- Oct 8, 2026 · Research paperA Closer Look at Agentic BBO: Benchmarking LLM Agents for Black-Box Optimization
- Oct 8, 2026 · Research paperInternalizer: Portable Context-to-Parameter Mapping for Very Large Language Models
- Oct 8, 2026 · Research paperMine Odyssey: Benchmarking Spatial Agentic Intelligence in the Wild
- Oct 7, 2026 · Research paperStoreBench: A Live-Commerce Environment for Evaluating and Training Autonomous Operator Agents
- Oct 7, 2026 · Research paperEvaluating Rubric Generation with Interventional Transfer
- Oct 6, 2026 · Research paperGeoNatureAgent (GNA): A Framework and Benchmark for Pre-Production Evaluation of Tool-Using Agents on Geospatial and Environmental Tasks
- Oct 6, 2026 · Research paperBeyond the Leaderboard: Multi-Dimensional Evaluation of Dense and Mixture-of-Experts Models for Automated Program Repair