Using llama.cpp release b9781 for quantization.
Original model: https://huggingface.co/deepreinforce-ai/Ornith-1.0-9B
Benchmarked by Infersec on:
Performance and quality results across different hardware configurations