DeepSeek-R1-Distill-Qwen-7B
Model size details
Family
Qwen
Open the family page when adjacent sizes change the decision.
Maker
Alibaba
Provider or project that owns the model lane.
Size
7.6B
Needs source review
VRAM
8 GB
Planning band for local quantized runs, not a benchmark guarantee.
Tasks
Chat, Reasoning
Primary local jobs this row should help decide.
Context
8K
Context length affects VRAM and practical throughput.
Runtime path
Ollama, LM Studio, llama.cpp
Confirm current runtime and quant support before choosing a setup.
Quantization
Q4, Q8, GGUF
The available quant changes memory use, quality, and throughput.
Decision summary and limitation
A small-to-mid text reasoning model for local exploration, with upstream lineage kept visible.
Limitation: Reasoning output can be long and slow; do not treat this entry as an accuracy, safety, or device-fit claim.
Source, freshness and evidence
Primary source: official-huggingface-deepseek
Freshness: Official model card checked 2026-08-06
Evidence: Official DeepSeek model card; MIT license with Qwen Apache-2.0 lineage noted.
Family variants
Adjacent model sizes in the same family.
| Model | Tasks | VRAM | Fit band |
|---|---|---|---|
| Qwen 2 5 Coder 7b | Chat, Code | 8 GB | Fast local default |
| DeepSeek-R1-Distill-Qwen-7B | Chat, Reasoning | 8 GB | Fast local default |
| Qwen 3 14B | Chat | 16 GB | Device fit requires local validation |