Pass IndexThe State of AISign in

Llama-3.1-Nemotron-70B-Instruct-HF

Llama-3.1-Nemotron-70B-Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM generated responses to user queries.

text → text · made by NVIDIA

$0.88$0.88per Mtok in / out
Together AI

Sold by 1 way

SellerLaneRate
Together AI aggregatorstandard$0.88per Mtok in$0.88per Mtok outraw.githubusercontent.com · read 2026-09-16

About

Llama-3.1-Nemotron-70B-Instruct-HF — a text model from NVIDIA, sold by one company from $0.88 in and $0.88 out per million tokens.

It takes text and returns text, with a context window of 16,384 tokens. It was published in October 2024. The catalogue files it under chat. Its weights are published under the llama3.1 licence, which attaches conditions the plain open licences do not. Only Together AI sells it, at $0.88 in and $0.88 out per million tokens.

Every current figure

Maker
NVIDIA
Register
model
Takes
text
Returns
text
Context
16,384 tokens
Longest answer
8,192 tokens
Published
October 2024
Parameters
71 billion · read from its own weights
Licence
llama3.1
Sellers
1
Maker's own price
not read
Price
$0.88 in and $0.88 out per million tokens — Together AI

Known as 1 name

nvidia/Llama-3.1-Nemotron-70B-Instruct-HF