Language
Nemotron 3 Nano Omni 30B-A3B (MoE, multimodal)
NVIDIA Nemotron 3 Nano Omni — 30B total / 3B active hybrid reasoning MoE, 256K context. Accepts audio, video, text, images and documents; output is text. Needs a llama.cpp-compatible backend: the vision path uses a separate mmproj file that Ollama does not load.
NVIDIA OpenSource verified
Specification
Model ID
nemotron-3-nano-omni-30b-a3bFamily
nemotron
Modality
Language
Parameters
30B (MoE, 3B active)
Context
262,144 tokens
Quantization
UD-Q4_K_XL
Weights
17.2 GB
Minimum RAM
25 GB
Source
Weights and provenance.
Registry license
NVIDIA Open
Access
Open — weights fetch without accepting additional terms.
The licence above is the registry's. The source repository either states nothing, states the catch-all “other”, or carries a licence link and base model that its own tag contradicts — a re-quantisation inherits the terms of the weights it derives from. The registry's licence is the one to rely on.
Serve it
# Pull the weights onto a node
tenzro model download nemotron-3-nano-omni-30b-a3b
# Serve it
tenzro model serve nemotron-3-nano-omni-30b-a3b
# Or query the registry entry over JSON-RPC
curl https://rpc.tenzro.xyz \
-X POST -H "Content-Type: application/json" \
-d '{"jsonrpc":"2.0","id":1,
"method":"tenzro_getModel",
"params":[{"model_id":"nemotron-3-nano-omni-30b-a3b"}]}'Same family