llama.cpp GGUF build (Q4_0 QAT) of Gemma 4 E4B IT. One projector file carries BOTH encoders - clip.has_vision_encoder and clip.has_audio_encoder - so the entry reads images and speech on the GGUF consensus engine at 6.1 GB instead of the 18 GB the torch path needs.
llama.cpp GGUF build (Q4_0 QAT) of Gemma 4 E4B IT. One projector file carries BOTH encoders - clip.has_vision_encoder and clip.has_audio_encoder - so the entry reads images and speech on the GGUF consensus engine at 6.1 GB instead of the 18 GB the torch path needs.