Tenzro
Language

NVIDIA Nemotron 3.5 Lightning 30B-A3B (MoE)

NVIDIA Nemotron 3.5 Lightning — 30B total / 3B active hybrid Mamba-2 + Attention MoE. Thinking + instruct, 256K native context extensible to ~1M. Single-file GGUF, single-node serving.
OpenMDW-1.1Source verified
Specification
Model ID
nemotron-3.5-lightning-30b-a3b
Family
nemotron-lightning
Modality
Language
Parameters
30B (MoE, 3B active)
Context
262,144 tokens
Quantization
UD-Q4_K_XL
Weights
23.7 GB
Minimum RAM
32 GB
Source

Weights and provenance.

Registry license
OpenMDW-1.1
Access
Open — weights fetch without accepting additional terms.

The licence above is the registry's. The source repository either states nothing, states the catch-all “other”, or carries a licence link and base model that its own tag contradicts — a re-quantisation inherits the terms of the weights it derives from. The registry's licence is the one to rely on.

Serve it
# Pull the weights onto a node
tenzro model download nemotron-3.5-lightning-30b-a3b

# Serve it
tenzro model serve nemotron-3.5-lightning-30b-a3b

# Or query the registry entry over JSON-RPC
curl https://rpc.tenzro.xyz \
  -X POST -H "Content-Type: application/json" \
  -d '{"jsonrpc":"2.0","id":1,
       "method":"tenzro_getModel",
       "params":[{"model_id":"nemotron-3.5-lightning-30b-a3b"}]}'
← All models