Developers

API reference.

Three APIs: the OpenAI-compatible inference endpoint your end users hit, the management REST API the console uses, and the agent enrollment protocol Hux speaks.

Inference (your gateway).

Standard OpenAI v1 surface. Hits your fleet directly, never our cloud.

POST https://your-gateway/v1/chat/completions
Authorization: Bearer sk_…
Content-Type: application/json

{
  "model": "llama-3.2-1b",
  "messages": [{"role":"user","content":"…"}],
  "stream": true,
  "max_tokens": 200
}

The model field accepts either the model's display name (e.g. google/gemma-3-1b-it) or its slugified routing key (google-gemma-3-1b-it). Hux normalises both. The /api/v1/* path also works for OpenAI client SDKs that prepend it.

Management REST API.

JSON over HTTPS, bearer-auth with your console session token. Multi-tenant, RLS-isolated by org.

Agent enrollment.

The protocol Hux uses to talk to Kan. You usually don't care about this — the installer handles it. Documented for operators rolling their own deployment automation.


Back to docs