Nemotron 3 Nano 30B A3B

nvidia/nemotron-3-nano-30b-a3b

输入 text · 输出 text · 上下文 262,144 · 最大输出 235,929

工具Schema推理缓存不用于训练
每千次请求的实际成本

成本 = p_in·(input − cached) + p_cache_read·cached + p_out·output + p_req;缓存 Token 是粘性端点已持有的前缀(part2b §9.3),因此标价最低并不总是单次请求最便宜。

供应商区域 · 量化上下文 / 最大输出输入 / 百万输出 / 百万缓存读取 / 写入加价每千次请求30 天可用率TTFT p50 / p95吞吐量最近 24 小时保真度数据能力状态
OpenRouterglobal · unknown262,144 / 235,929$0.053$0.210$0.032 / —+5.0%$0.190———无流量— 不用于训练 30 天模型默认活跃

快速开始

curl https://api.ai.ml/v1/chat/completions \
  -H "Authorization: Bearer $AIML_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"nvidia/nemotron-3-nano-30b-a3b","messages":[{"role":"user","content":"Say hello"}]}'
在控制台获取密钥

变更记录

  • new_model model nvidia/nemotron-3-nano-30b-a3b added (active); endpoint ep_openrouter_nemotron-3-nano-30b-a3b_global (nvidia/nemotron-3-nano-30b-a3b) added as active; price input 0.053 output 0.210 USD/M 2026/10/5