Make AI Fit
EN

DeepSeek-R1-Distill-Qwen-7B

QwenAlibaba7.6B8 GB8KFast local default

Model size details

Family

Qwen

Open the family page when adjacent sizes change the decision.

Maker

Alibaba

Provider or project that owns the model lane.

Size

7.6B

Needs source review

VRAM

8 GB

Planning band for local quantized runs, not a benchmark guarantee.

Tasks

Chat, Reasoning

Primary local jobs this row should help decide.

Context

8K

Context length affects VRAM and practical throughput.

Runtime path

Ollama, LM Studio, llama.cpp

Confirm current runtime and quant support before choosing a setup.

Quantization

Q4, Q8, GGUF

The available quant changes memory use, quality, and throughput.

Decision summary and limitation

A small-to-mid text reasoning model for local exploration, with upstream lineage kept visible.

Limitation: Reasoning output can be long and slow; do not treat this entry as an accuracy, safety, or device-fit claim.

Source, freshness and evidence

Primary source: official-huggingface-deepseek

Freshness: Official model card checked 2026-08-06

Evidence: Official DeepSeek model card; MIT license with Qwen Apache-2.0 lineage noted.

Family variants

Adjacent model sizes in the same family.

Open family
Model Tasks VRAM Fit band
Qwen 2 5 Coder 7b Chat, Code 8 GB Fast local default
DeepSeek-R1-Distill-Qwen-7B Chat, Reasoning 8 GB Fast local default
Qwen 3 14B Chat 16 GB Device fit requires local validation