Qwen3-8B-GGUF

Qwen3-8B-GGUF

VRAM needed

16 GB

Fits 16 GB+ consumer GPUs

AuthorQwen
Parameters8B
QuantizationQ4_K_M
FormatGGUF
Licenseapache-2.0
File size4.68 GB
View on HuggingFace

Overview

The official Qwen3 8B dense model - a reliable, well-rounded small model for chat, light coding, and tool use on 16 GB hardware.

Recommended hardware

≈ 16 GB VRAM to run this 8B model at the Q4_K_M quantization.

Fits 16 GB+ consumer GPUs (e.g. RTX 4060 Ti 16GB, RTX 4080)

Links