Base URL
One host serves everything:https://platform.omnia-voice.com/api/v1 keep working; it’s an alias,
not a migration.
Everything authenticates with the same workspace API key
(Authorization: Bearer sk_sovereign_...). Create one in the
dashboard under API keys.
Your first request
usage
object reflects the exact tokens deducted from your wallet. See the
Quickstart for the SDK equivalents.
Start here
How Omnia works
The workflow end to end: capture, measure, improve, own.
Quickstart
From an API key to your first response in a few minutes.
Authentication
How keys work, the
sk_sovereign_ format, and scoping.API reference
Every endpoint and field, interactive, generated from the API.
The gateway
How a request is authenticated, routed, metered, and settled.
Measure and improve
Grade your traffic
Record pass/fail verdicts on real traffic: the ground truth every judge
is measured against.
Judges & calibration
Calibrate every judge against your grades: TPR, TNR, Cohen’s κ, and a
trust badge derived from confidence intervals.
Comparisons
Test whether a cheaper model holds up on your own prompts, with pass rates
corrected for the judge’s measured error.
Self-improvement
Use a calibrated judge as a training reward, and prove the result on
held-out grades.
Run at scale
Inference
Chat, streaming, tool calling, structured output, embeddings, and vision,
all OpenAI-compatible.
Dedicated endpoints
Reserve private GPU capacity for a model, billed per GPU-hour while running
instead of per token.
Fine-tuning
Train a model on your own data, then serve it on a dedicated endpoint and
call it like any other model.
Billing
Prepaid wallet, per-token and per-GPU-hour metering, holds, and auto-reload.