Nemotron 3 Nano Omni 30B A3B Reasoning
Chatnvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16
Nemotron 3 Nano Omni 30B A3B Reasoning on FlexAI: NVIDIA LLM (Multimodal), NVIDIA Open Model Agreement license, available as a dedicated endpoint on FlexAI or your own infrastructure.
Pricing
Input
$0.2 / M tokens
Output
$0.8 / M tokens
Cached input
$0.03 / M tokens
Context
256K tokens
API endpoint
/v1/chat/completions
Compatibility
OpenAI
Parameters
33.02B total / 3.0B active
License
NVIDIA Open Model Agreement
Hardware
H100
Quantization
BF16
Nemotron 3 Nano Omni 30B A3B Reasoning runs as a dedicated endpoint, provisioned per customer on FlexAI's infrastructure or your own, not served through the shared Token Factory API.