Pass IndexThe State of AISign in

DeepSeek-R1-Distill-QWEN-7B

DeepSeek-R1-Distill-Qwen-7B is a 7B model built on Qwen2.5-Math-7B and fine-tuned on 800,000 samples generated by DeepSeek-R1, distilling the larger model's reasoning patterns into a compact architecture.

text → text · made by DeepSeek

$0.2$0.2per Mtok in / out
Fireworks AI · 3 sellers

Sold by 3 ways

SellerLaneRate
Hugging Face Inference Providers aggregatornscale$0.15per Mtok in$0.15per Mtok outrouter.huggingface.co · read 2026-08-25
Fireworks AI aggregatorstandard$0.2per Mtok in$0.2per Mtok outraw.githubusercontent.com · read 2026-09-16
Nscale aggregatorstandard$0.2per Mtok in$0.2per Mtok outraw.githubusercontent.com · read 2026-09-16

About

DeepSeek-R1-Distill-QWEN-7B — a text model from DeepSeek, sold by 3 companies from $0.2 in and $0.2 out per million tokens.

It takes text and returns text, with a context window of 32,768 tokens. It was published in 2025. Its sellers say it can reason step by step and call a tool. The catalogue files it under chat. Three companies sell it. The cheapest is $0.2 in and $0.2 out per million tokens at Fireworks AI. Beside the standard rate there is a separately routed lane.

Every current figure

Maker
DeepSeek
Register
model
Takes
text
Returns
text
Context
32,768 tokens
Longest answer
16,384 tokens
Published
2025
Parameters
7 billion · read from its own name
Licence
not read
Sellers
3
Maker's own price
not read
Price
$0.2 in and $0.2 out per million tokens — Fireworks AI

Known as 1 name

deepseek-ai/DeepSeek-R1-Distill-Qwen-7B