Pass IndexThe State of AISign in

Llama-3.1-Nemotron-70B-Instruct

Llama-3.1-Nemotron-70B-Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM generated responses to user queries.

text → text · made by NVIDIA

$1.2$1.2per Mtok in / out
DeepInfra

Sold by 1 way

SellerLaneRate
DeepInfra aggregatorstandard$1.2per Mtok in$1.2per Mtok outapi.deepinfra.com · read 2026-08-25

About

Llama-3.1-Nemotron-70B-Instruct — a text model from NVIDIA, sold by one company from $1.2 in and $1.2 out per million tokens.

It takes text and returns text, with a context window of 131,072 tokens. It was published in April 2025. Its sellers say it can call a tool. The catalogue files it under chat. Only DeepInfra sells it, at $1.2 in and $1.2 out per million tokens.

Every current figure

Maker
NVIDIA
Register
model
Takes
text
Returns
text
Context
131,072 tokens
Longest answer
8,192 tokens
Published
April 2025
Parameters
70 billion · read from its own name
Licence
not read
Sellers
1
Maker's own price
not read
Price
$1.2 in and $1.2 out per million tokens — DeepInfra

Known as 1 name

nvidia/Llama-3.1-Nemotron-70B-Instruct