VRAM needed
≈ 24 GB
Fits 24 GB+ consumer GPUs
Qwen3-Coder Next 80B at an aggressive 2-bit quant - frontier-class coding quality that still fits a 24 GB GPU via heavy quantization.
≈ 24 GB VRAM to run this 80B model at the UD-IQ2_XXS quantization.
Fits 24 GB+ consumer GPUs (e.g. RTX 3090, RTX 4090)