whichllm — Browse and compare AI model specs and pricing

Requesty

nvidia-nemotron-3-ultra model ID, context window & pricing

nemotron

Quick facts

Model ID nvidia-nemotron-3-ultra
Source Requesty
Context Window 262144
Pricing $0.50 input / $2.50 output per 1M tokens
Capabilities reasoning

Model overview

nvidia-nemotron-3-ultra is an AI model from Requesty with 262144 token context window and text input support.

Published pricing is $0.50 input and $2.50 output per 1M tokens.

  • Workloads that use text inputs with text outputs.
  • Reasoning-heavy prompts where stepwise problem solving matters.
Model ID nvidia-nemotron-3-ultra
Provider Requesty
Family nemotron
Status -
Knowledge Cutoff -
Release Date 2026-06-23
Input Modalities text
Output Modalities text
Context Window 262144
Input Limit -
Output Limit 131072
Tool Calling No
Reasoning Yes
Structured Output No
Temperature Control -
Open Weights No
Input Cost / 1M tokens $0.50
Output Cost / 1M tokens $2.50
Reasoning Cost / 1M tokens -
Cache Read Cost / 1M tokens -
Cache Write Cost / 1M tokens -