Contributors › julianmb's local-LLM benchmark configs

julianmb's local-LLM benchmark configs

Every tracked benchmark config measured by julianmb: model, quant, backend, tokens per second, source-linked and trust-tiered.

Snapshot 2026-09-16 · 158 configs · 75 models. Every number is generated from the tracker snapshot and linked to where it was measured; nothing here is typed by hand.

10 configs · 4 models · 211 stars · 76.9 tok/s best single-stream · julianmb on github.

Configs measured by julianmb

ModelQuantBackendDecodeContextTrustSource
Ornith 1.5 35B-A3BFP476.9 tok/sOwner-submittedsource
Ornith 1.5 35B-A3BQ4_K_M71.6 tok/sOwner-submittedsource
Qwen 3.8 27BFP433.8 tok/sOwner-submittedsource
DeepSeek V4 Flash 284BIQ2_XXS32 tok/sOwner-submittedsource
Qwen 3.8 27BFP426.9 tok/s32768Owner-submittedsource
Qwen3.8-27B-DFlash2Q4_K_Mvulkan21.2 tok/s (prefill 211)32768Owner-submittedsource
Qwen 3.8 27BFP8llama.cpp19 tok/sOwner-submittedsource
Qwen 3.8 27BQ4_K_M12.4 tok/sOwner-submittedsource
Qwen 3.8 27BQ4_K_Mllama.cpp12.3 tok/s32768Owner-submittedsource
Qwen 3.8 27BFP16llama.cpp5 tok/sOwner-submittedsource

More: all configs · hardware · methodology · llms.txt · llms-full.txt · API (OpenAPI) · JSON snapshot. Built by Altronis, private on-prem AI, Singapore.