gemma-4-26B-A4B-it · Q4_K_M
L1 ReproducedApple M2 Ultra · llama.cpp
GitHub
Yesterday
79.11 tok/s
Decode
1,550 tok/s
Prefill
— s
TTFT
— GB
VRAM
Config (latest run)
macOS 26.5.1 Q4_K_M llama.cpp b11050
Sources: GitHub · 1 runs · 1 independent sources