NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF

NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF

VRAM needed

32 GB

Needs a 32 GB+ GPU (or multi-GPU)

Authorunsloth
Parameters30B
QuantizationUD-Q4_K_M
FormatGGUF
Licenseopenmdw-1.1
File size23.53 GB
View on HuggingFace

Overview

NVIDIA's Nemotron 3.5 Lightning at Unsloth's dynamic 4-bit - a Mamba-2 hybrid MoE with 3B active parameters, toggleable reasoning, and strong instruction following (IFBench 71.9). Needs a 32 GB GPU.

Recommended hardware

≈ 32 GB VRAM to run this 30B model at the UD-Q4_K_M quantization.

Needs a 32 GB+ GPU (e.g. RTX 5090, or multi-GPU / 2x24 GB)

Links