API - TokenMark OpenAI-Compatible Local Inference

Access local LLM inference via an OpenAI-compatible API. Run models in your browser with WebGPU acceleration. Free tier available.

Snapshot 2026-09-30 · 254 configs · 116 models. Every number is generated from the tracker snapshot and linked to where it was measured; nothing here is typed by hand.

More: all configs · hardware · methodology · llms.txt · llms-full.txt · API (OpenAPI) · JSON snapshot. Built by Altronis, private on-prem AI, Singapore.