TinyR1-32B-Preview
We applied supervised fine-tuning (SFT) to Deepseek-R1-Distill-Qwen-32B across three target domains—Mathematics, Code, and Science — using the 360-LLaMA-Factory training framework to produce three domain-specific models.
text → text · made by qihoo360
Sold by
Nobody in the catalogue publishes a price for this yet.
About
TinyR1-32B-Preview — a text model from qihoo360.
It takes text and returns text. It was published in February 2025. The catalogue files it under chat. Its weights are published under APACHE-2.0, so you may run it on your own machine, or buy it from whoever serves it cheapest.
Every current figure
- Maker
- qihoo360
- Register
- model
- Takes
- text
- Returns
- text
- Published
- February 2025
- Parameters
- 33 billion · read from its own weights
- Licence
- apache-2.0
Known as 1 name
qihoo360/TinyR1-32B-Preview