Contributors › theNetworkChuck's local-LLM benchmark configs
theNetworkChuck's local-LLM benchmark configs
Every tracked benchmark config measured by theNetworkChuck: model, quant, backend, tokens per second, source-linked and trust-tiered.
Snapshot 2026-09-30 · 254 configs · 116 models. Every number is generated from the tracker snapshot and linked to where it was measured; nothing here is typed by hand.
3 configs · 3 models · 7 stars · 109.1 tok/s best single-stream · theNetworkChuck on github.
Configs measured by theNetworkChuck
| Model | Quant | Backend | Decode | Context | Trust | Source |
|---|---|---|---|---|---|---|
| Qwen3.6 35B-A3B | MLX-4BIT | mlx | 109.1 tok/s (prefill 6012) | 32000 | Extracted, source-linked | source |
| Qwen3 14B | MLX-4BIT | mlx | 55.4 tok/s (prefill 2278) | 32000 | Extracted, source-linked | source |
| Qwen3.8 27B | MLX-4BIT | mlx | 42.6 tok/s (prefill 1526) | 32000 | Extracted, source-linked | source |
More: all configs · hardware · methodology · llms.txt · llms-full.txt · API (OpenAPI) · JSON snapshot. Built by Altronis, private on-prem AI, Singapore.