跳过主要内容
NVNVIDIA/ Nemotron

Nemotron 3 Ultra 550B A55B

nvidia/nemotron-3-ultra-550b-a55b

Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy

模型信息

发布日期
2026-06-04
上下文窗口
1M
最大输出
128K
权重
开放

模态

输入
文本
输出
文本

模型身份

版本
nemotron-3-ultra-550b-a55b
别名
Nemotron 3 Ultra 550B A55B

能力

开放权重推理温度参数工具调用

官方价格

输入 $0.5/M · 输出 $2.5/M
缓存读取 $0.15/M

提供商

提供商提供商模型 ID协议上下文最大输出输入输出能力备注价格流式输出
Nvidianvidia/nemotron-3-ultra-550b-a55bcustom1M128K
推理工具调用结构化输出温度参数
-
输入 $0.5/M · 输出 $2.5/M
缓存读取 $0.15/M
Kenarinemotron-3-ultra-550b-a55bopenai_chat1M128K
推理工具调用温度参数
-
输入 $0/M · 输出 $0/M
UnoRouternemotron-3-ultra-550b-a55b:freeopenai_chat1M128K
推理工具调用温度参数
-
输入 $0/M · 输出 $0/M
Together AInvidia/nemotron-3-ultra-550b-a55bopenai_chat1M128K
推理工具调用结构化输出温度参数
-
输入 $0.6/M · 输出 $3.6/M
缓存读取 $0.2/M
Vercel AI Gatewaynvidia/nemotron-3-ultra-550b-a55bopenai_chat1M128K
推理工具调用温度参数
-
输入 $0.6/M · 输出 $2.4/M
缓存读取 $0.12/M
OpenRouternvidia/nemotron-3-ultra-550b-a55bopenai_chat1M128K
推理工具调用结构化输出温度参数
可选思考强度: high, medium, 默认推理强度: high, 思考支持 token 预算, Hugging Face: nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16
输入 $0.6/M · 输出 $3.6/M
缓存读取 $0.2/M
OpenRouternvidia/nemotron-3-ultra-550b-a55b:freeopenai_chat1M128K
推理工具调用温度参数
可选思考强度: high, medium, 默认推理强度: high, 思考支持 token 预算, Hugging Face: nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16
输入 $0/M · 输出 $0/M
NanoGPTnvidia/nemotron-3-ultra-550b-a55bopenai_chat1M128K
推理工具调用温度参数
-
输入 $0.5/M · 输出 $2.5/M
缓存读取 $0.25/M
Kilo Gatewaynvidia/nemotron-3-ultra-550b-a55b:freeopenai_chat1M128K
推理工具调用温度参数
-
输入 $0/M · 输出 $0/M
Kilo Gatewaynvidia/nemotron-3-ultra-550b-a55bopenai_chat1M128K
推理工具调用结构化输出温度参数
-
输入 $0.5/M · 输出 $2.2/M
缓存读取 $0.1/M

常见问题