Skip to content

    Nemotron 3 Nano Omni 30B A3B Reasoning

    Chat

    nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16

    Nemotron 3 Nano Omni 30B A3B Reasoning on FlexAI: NVIDIA LLM (Multimodal), NVIDIA Open Model Agreement license, available as a dedicated endpoint on FlexAI or your own infrastructure.

    Pricing

    Input

    $0.2 / M tokens

    Output

    $0.8 / M tokens

    Cached input

    $0.03 / M tokens

    Context

    256K tokens

    API endpoint

    /v1/chat/completions

    Compatibility

    OpenAI

    Parameters

    33.02B total / 3.0B active

    License

    NVIDIA Open Model Agreement

    Hardware

    H100

    Quantization

    BF16

    Nemotron 3 Nano Omni 30B A3B Reasoning runs as a dedicated endpoint, provisioned per customer on FlexAI's infrastructure or your own, not served through the shared Token Factory API.