Llama-3.3-Nemotron-Super-49B-v1.5
Llama-3.3-Nemotron-Super-49B-v1.5 is a large language model (LLM) optimized for advanced reasoning, conversational interactions, retrieval-augmented generation (RAG), and tool-calling tasks.
text → text · made by NVIDIA
$0.4→$0.4per Mtok in / out
DeepInfra
Sold by 2 ways
| Seller | Lane | Rate |
|---|---|---|
| DeepInfra aggregator | standard | $0.4per Mtok in$0.4per Mtok outapi.deepinfra.com · read 2026-08-25 |
| ElectronHub 0 ways | $1per Mtok in$2per Mtok out | |
| ElectronHub aggregatornot on the seller's list today | standard | $1per Mtok in$2per Mtok outapi.electronhub.ai · read 2026-08-29 |
About
Llama-3.3-Nemotron-Super-49B-v1.5 — a text model from NVIDIA, sold by one company from $0.4 in and $0.4 out per million tokens.
It takes text and returns text, with a context window of 131,072 tokens. It was published in July 2025. Its sellers say it can reason step by step and call a tool. The catalogue files it under search. Only DeepInfra sells it, at $0.4 in and $0.4 out per million tokens.
Every current figure
- Maker
- NVIDIA
- Register
- model
- Takes
- text
- Returns
- text
- Context
- 131,072 tokens
- Published
- July 2025
- Parameters
- 49 billion · read from its own name
- Sellers
- 1
- Maker's own price
- not read
- Price
- $0.4 in and $0.4 out per million tokens — DeepInfra
Known as 1 name
nvidia/Llama-3.3-Nemotron-Super-49B-v1.5