The same weights, for less

Cut the token bill.Keep the performance.

Run open-weight models for less, without changing how you work. On GLM-5.2, Draupnir delivers the same model at lower output-token cost than the standard hosted price — same weights, no quantization tricks.

Illustrative. GLM-5.2 public hosted list price vs Draupnir list price, August 2026.

01 — Models & pricing

Models & pricing

The same open weights, below the public hosted rate. Prices are per million tokens.

GLM-5.2 model availability and pricing
ModelContextInput $/MOutput $/MPublic output $/MStatus
GLM-5.2500kPendingPending$3.85Live

Pricing shared on API access.

Weights on Hugging Face

Benchmark figures and third-party verification pending.

02Coverage policy

New open-weight models, served within days of release.

03API surface

OpenAI-compatible.
Change the base URL and ship.

Get an API key

Access details shared after review.

02 — Who it is for

Two ways the token bill hits you

01 / Serving teams

Teams serving open weights

Your margin is the gap between what a token costs you to serve and what you charge for it. Draupnir widens it without changing the model.

02 / Routers

Gateways and routers

You select providers on price and latency at a measurable target. Draupnir earns the route by occupying the better point on that surface.