1. Application Performance Monitoring intermediate Multiple choice

    When does APM tell you something that distributed tracing and metrics cannot?

    2 min answer apmtracingprofilingtooling
  2. Business Metrics intermediate

    A marketplace's technical dashboards are all green during a checkout failure that costs significant revenue. What kind of monitoring was missing?

    2 min answer meeshobusiness-metricsdetectionfunnels
  3. Business Metrics intermediate

    A marketplace's technical metrics are all healthy during an incident in which sellers cannot list items. Why did technical monitoring miss it, and what should be measured?

    2 min answer business-metricsdetectionsilent-failureetsy
  4. Business Metrics advanced

    Customers reported an outage 40 minutes before your monitoring did. How do you close that gap?

    2 min answer detectionbusiness-metricssyntheticsclient-side
  5. Cardinality advanced Multiple choice

    A live-streaming platform of Twitch's shape needs usable latency quantiles for 300 API endpoints whose responses span 2 ms to 90 seconds. The current histograms use 12 fixed buckets topping out at 10 seconds and every p99 above that reads as the overflow bucket. Which change fits the problem?

    3 min answer cardinalityhistogramsopentelemetryquantiles
  6. Cardinality advanced

    A single deploy took down your monitoring platform. What happened, and how do you prevent a recurrence?

    2 min answer cardinalitymetricscostguardrails
  7. Cardinality advanced

    A team wants to debug failures nobody predicted, using wide high-cardinality events rather than pre-aggregated metrics. How does the storage design differ, and how must sampling preserve rare errors?

    3 min answer honeycombwide-eventscardinalitysampling
  8. Cardinality advanced

    An observability platform ingests billions of telemetry events daily, and some customers attach dimensions with unbounded distinct values. How should ingestion, storage, indexing and query be designed so cost and latency stay manageable?

    2 min answer cardinalitytime-seriesingestionmulti-tenancy
  9. Cardinality advanced

    An observability platform ingests enormous telemetry volume, but a small number of dimensions create extreme cardinality. How should ingestion, aggregation, indexing, sampling, retention and storage tiers be designed so cost and query performance stay predictable?

    2 min answer grafanacardinalitytelemetry-costretention
  10. Correlation IDs beginner Multiple choice

    A customer reports an error at 14:32 yesterday. What must have been built for the investigation to take two minutes rather than two hours?

    2 min answer correlation-idssupportobservabilitydebugging
  11. Correlation IDs intermediate

    A logistics platform's request crosses synchronous services, message queues, scheduled batch jobs and third-party callbacks. Tracing works within services and breaks between them. What is missing?

    2 min answer delhiverycorrelationasynctracing
  12. Correlation IDs intermediate

    An API platform's requests trigger asynchronous work, webhook deliveries and retries, sometimes hours later. How should correlation identifiers be designed so a customer question can be answered end to end?

    2 min answer correlation-idscausalityasyncsupport