VRAM needed
≈ 16 GB
Fits 16 GB+ consumer GPUs
Compact Ornith 1.0 dense model - fast and frugal, ideal for budget GPUs (16 GB) and latency-sensitive serving. A great starter model with a long context window.
≈ 16 GB VRAM to run this 9B model at the Q6_K quantization.
Fits 16 GB+ consumer GPUs (e.g. RTX 4060 Ti 16GB, RTX 4080)