AI Executive Office — CXO Assistant Platform · View 23 of 30 · 6 · Operations
What is measured
- Latency, tokens and cost per turn per tenant; tool failure rate per source; grounding rate and abstention rate; data freshness per source against its contract
- Recommendation acceptance is treated as a platform metric. If executives stop acting on what the platform surfaces, that is an outage of the product
- Traces follow the OpenTelemetry GenAI conventions so a turn can be read end to end: prompt, plan, each tool call, model version, grounding verdict
The constraint
- The platform SRE must diagnose a degraded tenant without reading tenant content. Telemetry carries identifiers, timings and verdicts — never prompts, retrieved passages or business values. Content-level traces stay inside the tenant boundary and are readable only by the tenant's own governance role
- Stale data is shown in the answer, never handled silently
Numbers
- 90 days hot in Log Analytics, 2 years archived
- Cost per interaction is a tracked unit-economic metric with a per-tenant budget and an alert
- SLO burn alerts on p95 answer latency and on grounding rate over a rolling 24 hours