VRAM needed
≈ 4 GB
Fits 4 GB+ consumer GPUs, or CPU inference
The native BF16 checkpoint of Liquid's LFM2.5 230M for vLLM - a 230M edge model with a 32K context window, suited to data extraction and lightweight on-device agent pipelines.
≈ 4 GB VRAM to run this 230M model at the BF16 quantization.
Fits 4 GB+ consumer GPUs (e.g. GTX 1650), or CPU inference