DeepSeek-R1-Distill-QWEN-7B
DeepSeek-R1-Distill-Qwen-7B is a 7B model built on Qwen2.5-Math-7B and fine-tuned on 800,000 samples generated by DeepSeek-R1, distilling the larger model's reasoning patterns into a compact architecture.
text → text · made by DeepSeek
$0.2→$0.2per Mtok in / out
Fireworks AI · 3 sellers
Sold by 3 ways
| Seller | Lane | Rate |
|---|---|---|
| Hugging Face Inference Providers aggregator | nscale | $0.15per Mtok in$0.15per Mtok outrouter.huggingface.co · read 2026-08-25 |
| Fireworks AI aggregator | standard | $0.2per Mtok in$0.2per Mtok outraw.githubusercontent.com · read 2026-09-16 |
| Nscale aggregator | standard | $0.2per Mtok in$0.2per Mtok outraw.githubusercontent.com · read 2026-09-16 |
About
DeepSeek-R1-Distill-QWEN-7B — a text model from DeepSeek, sold by 3 companies from $0.2 in and $0.2 out per million tokens.
It takes text and returns text, with a context window of 32,768 tokens. It was published in 2025. Its sellers say it can reason step by step and call a tool. The catalogue files it under chat. Three companies sell it. The cheapest is $0.2 in and $0.2 out per million tokens at Fireworks AI. Beside the standard rate there is a separately routed lane.
Every current figure
- Maker
- DeepSeek
- Register
- model
- Takes
- text
- Returns
- text
- Context
- 32,768 tokens
- Longest answer
- 16,384 tokens
- Published
- 2025
- Parameters
- 7 billion · read from its own name
- Licence
- not read
- Sellers
- 3
- Maker's own price
- not read
- Price
- $0.2 in and $0.2 out per million tokens — Fireworks AI
Known as 1 name
deepseek-ai/DeepSeek-R1-Distill-Qwen-7B