Configs › gemma-3-27b on Strix Halo, DGX Spark and Mac: measured tok/s
gemma-3-27b on Strix Halo, DGX Spark and Mac: measured tok/s
Every tracked benchmark config for gemma-3-27b on local hardware: quant, backend, context, tokens per second, source-linked and trust-tiered.
Snapshot 2026-09-30 · 254 configs · 116 models. Every number is generated from the tracker snapshot and linked to where it was measured; nothing here is typed by hand.
Measured configs for gemma-3-27b
| Model | Quant | Backend | Decode | Context | Trust | Source |
|---|---|---|---|---|---|---|
| gemma-3-27b | MLX-4BIT | mlx | 31.8 tok/s (prefill 854) | 512 | Extracted, source-linked | source |
| gemma-3-27b | Q4_K_M | llama.cpp | 29.9 tok/s (prefill 751) | 512 | Extracted, source-linked | source |
More: all configs · hardware · methodology · llms.txt · llms-full.txt · API (OpenAPI) · JSON snapshot. Built by Altronis, private on-prem AI, Singapore.