NVIDIA: Nemotron 3 Nano 30B A3B (free)

nvidia/nemotron-3-nano-30b-a3b:free

Description

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully open with open-weights, datasets and recipes so developers can easily customize, optimize, and deploy the model on their infrastructure for maximum privacy and security.

How this model compares

Overall covers the full catalog. By plan covers only models available on that tier (same rules as available models in your list). Position on min–average–max. Prices use the higher of prompt or completion per token, shown per 1M tokens.

Price (per 1M tokens)

Min
Max
This model
339 models in this groupPrice (per 1M tokens)
Min
$0.04
Avg
$12.395447
Max
$750.00
This model: $0.00 / 1M tokens

Context length (tokens)

Min
Max
This model
339 models in this groupContext length (tokens)
Min
4,095 tokens
Avg
379,884.782 tokens
Max
10,000,000 tokens
This model: 256,000 tokens

Capabilities

Text → TextContext: 256,000 tokens
Input:
Text
Output:
Text
    NVIDIA: Nemotron 3 Nano 30B A3B (free)