Changelog v4.4 fix 6 May 2026

Token accounting corrected for mixture-of-experts models

We were counting tokens against the total parameter count rather than the active experts, which overcharged Mixtral and DeepSeek V3 requests by roughly eight per cent. Affected accounts have been credited automatically; nobody needs to file anything.

Put the model next to the user.

One command, thirty-one regions, and an invoice that matches what you actually served.