All models

Nemotron 3 Ultra 550B A55B

NVIDIA
fireworks-ai/nemotron-3-ultra-nvfp4

Released Jun 2, 2026

providers
1
context
262.1K
max out
32.7K
input /M
$0.60
output /M
$2.40
cached input /M
$0.119
cache write /M
$0.60
reasoning /M
$2.40
also answers tonvidia-nemotron-3-ultra-550b-a55bnvidia/Nemotron-3-Ultra-550b-a55bnvidia/nemotron-3-ultra-550b-a55b-20260604

About

open weights
yes
best for
extractioncodingcompute

Capabilities

What the model accepts and returns, where a probe through our own gateway proved it.

  • streaming
  • tools
  • json mode
  • vision
  • audio