llama.cpp GGUF build (Q4_K_M) of Qwen3.5-9B - one artifact runs on Metal, CUDA and CPU; the build the GGUF consensus engine was proven on.
llama.cpp GGUF build (Q4_K_M) of Qwen3.5-9B - one artifact runs on Metal, CUDA and CPU; the build the GGUF consensus engine was proven on.