Skip to main content
The first command previews the llama-quantize command. --run is required to execute it, and the output directory is created automatically. Install a llama.cpp build exposing llama-quantize before running conversions.