Llama-3.1-Nemotron-70B-Instruct-HF
Llama-3.1-Nemotron-70B-Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM generated responses to user queries.
text → text · made by NVIDIA
$0.88→$0.88per Mtok in / out
Together AI
Sold by 1 way
| Seller | Lane | Rate |
|---|---|---|
| Together AI aggregator | standard | $0.88per Mtok in$0.88per Mtok outraw.githubusercontent.com · read 2026-09-16 |
About
Llama-3.1-Nemotron-70B-Instruct-HF — a text model from NVIDIA, sold by one company from $0.88 in and $0.88 out per million tokens.
It takes text and returns text, with a context window of 16,384 tokens. It was published in October 2024. The catalogue files it under chat. Its weights are published under the llama3.1 licence, which attaches conditions the plain open licences do not. Only Together AI sells it, at $0.88 in and $0.88 out per million tokens.
Every current figure
- Maker
- NVIDIA
- Register
- model
- Takes
- text
- Returns
- text
- Context
- 16,384 tokens
- Longest answer
- 8,192 tokens
- Published
- October 2024
- Parameters
- 71 billion · read from its own weights
- Licence
- llama3.1
- Sellers
- 1
- Maker's own price
- not read
- Price
- $0.88 in and $0.88 out per million tokens — Together AI
Known as 1 name
nvidia/Llama-3.1-Nemotron-70B-Instruct-HF