Every xAI Grok call, measured and priced.
xAI's Grok models run behind an OpenAI-compatible API; Obsivara tells you what each call costs and how reliably it runs. Every completion and tool call traced with tokens, latency, and dollars attached, priced per Grok model and rolled up per workflow — and compared against your other providers on real traffic, so routing a workload to Grok is a measured decision. Instrument via the Obsivara SDK or OpenTelemetry, or POST a usage event, out of the request path.
Grok (xAI), fully observable.
Per-call cost and token accounting
Input and output tokens and the dollar cost of every request, priced per Grok model and attributed to the workflow that made it.
Latency and error monitoring
P50/P95/P99 latency per model, rate-limit tracking, and error classification — so you know when Grok is slow versus when your prompts are.
Model comparison on live traffic
See what moving a workload between Grok tiers, or between Grok and another provider, would save — measured on your actual requests.
Usage anomaly detection
Token spikes, runaway retries, and prompt-size creep flagged before they land on your invoice.
Three steps. No code changes.
Add your Grok usage source
Instrument your app with the Obsivara SDK or an OpenTelemetry GenAI exporter, or POST one usage event per request. xAI's OpenAI-compatible API is captured without special handling. Your app keeps calling Grok directly — Obsivara records each call out-of-band, so nothing routes through us and no latency is added.
Obsivara maps your usage
Within minutes you get a live inventory of every Grok model in use, the workflows calling it, and what each costs.
Set budgets and alerts
Spend thresholds, latency SLOs, and error-rate alerts per workflow. Obsivara watches from there.
Signals tracked out of the box.
- →Tokens per request (input / output)
- →Cost per call, workflow, and day
- →Latency P50 / P95 / P99 by model
- →Rate-limit and quota errors
- →Model and parameter drift between deploys
- →Prompt size creep over time
Common questions.
xAI's Grok API is OpenAI-compatible — does Obsivara just work?
Yes. Instrument with the Obsivara SDK, OpenTelemetry, or usage events; the OpenAI-compatible request and response shape is captured without special configuration.
Does Obsivara need my xAI API key?
No. Usage is captured from events your app sends via the SDK, OpenTelemetry, or a webhook; Obsivara never needs your production keys or calls Grok on your behalf.
Will this add latency to my Grok calls?
No. Obsivara never sits in your request path — your app calls Grok directly and reports usage afterward, so there's no proxy and no added latency.