Llama-3.1-Nemotron-70B-Instruct
Llama-3.1-Nemotron-70B-Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM generated responses to user queries.
text → text · made by NVIDIA
$1.2→$1.2per Mtok in / out
DeepInfra
Sold by 1 way
| Seller | Lane | Rate |
|---|---|---|
| DeepInfra aggregator | standard | $1.2per Mtok in$1.2per Mtok outapi.deepinfra.com · read 2026-08-25 |
About
Llama-3.1-Nemotron-70B-Instruct — a text model from NVIDIA, sold by one company from $1.2 in and $1.2 out per million tokens.
It takes text and returns text, with a context window of 131,072 tokens. It was published in April 2025. Its sellers say it can call a tool. The catalogue files it under chat. Only DeepInfra sells it, at $1.2 in and $1.2 out per million tokens.
Every current figure
- Maker
- NVIDIA
- Register
- model
- Takes
- text
- Returns
- text
- Context
- 131,072 tokens
- Longest answer
- 8,192 tokens
- Published
- April 2025
- Parameters
- 70 billion · read from its own name
- Licence
- not read
- Sellers
- 1
- Maker's own price
- not read
- Price
- $1.2 in and $1.2 out per million tokens — DeepInfra
Known as 1 name
nvidia/Llama-3.1-Nemotron-70B-Instruct