Quiz
2667 questions of the kind that actually get asked — in interviews, in architecture review boards, and by the person who has to run the thing at 3 AM. Every answer states the trade-off rather than the slogan, and says when the obvious choice is the wrong one.
All areas2667
Architecture Fundamentals81
Distributed Systems101
Data Architecture90
Cloud Architecture87
Networking86
API & Integration Architecture78
Reliability & Resilience88
Observability81
Performance & Capacity Engineering90
Security Architecture95
Cost Architecture & FinOps92
Business Architecture93
Architecture Communication91
Enterprise Architecture91
Legacy Modernization92
AI-Era Architecture86
Software Architecture & Engineering84
Architecture Patterns84
Architecture Decision-Making91
The Architect's Meta-Skills92
Delivery & Release Engineering93
Platform Engineering & Developer Experience92
Testing & Quality Architecture90
Data Platform Architecture88
Streaming & Real-Time Data93
Data Governance & Semantics81
Frontend & Experience Architecture91
Edge, Mobile & IoT88
Regulatory & Data Protection Architecture90
Assurance, Audit & Model Risk88
36 questions in Observability.
-
Metrics intermediate
Etsy released StatsD in 2011: a small daemon that receives metric samples from applications over UDP, aggregates them in memory, and flushes to the metrics backend on a fixed interval. Why was UDP the right choice for that design, and what did it make permanently impossible?
3 min answer etsystatsdmetricsudp -
Metrics intermediate
Every service dashboard shows healthy metrics while users cannot complete a purchase. How is that possible and what would have detected it?
2 min answer metricsbusiness-metricsdetectionblind-spots -
Observability intermediate
A new service goes live next week. What observability must exist on day one?
2 min answer observabilitylaunchoperations -
Observability intermediate
You are designing a new service. What must be in place before it goes live so that whoever is paged at 3 AM can diagnose it without you?
2 min answer observabilityoncallalertinglogging -
OpenTelemetry intermediate
A platform with services in five languages and three monitoring vendors considers adopting OpenTelemetry. What does it solve, and what is the realistic migration cost?
2 min answer opentelemetrystandardisationvendor-lock-inmigration -
OpenTelemetry intermediate
An organisation with several existing telemetry systems considers adopting OpenTelemetry. What does it actually solve, and what does adopting it not fix?
2 min answer segmentopentelemetrystandardsvendor-lock-in -
SLO Monitoring intermediate Multiple choice
A checkout API has a 99.9% availability SLO and the team must decide where the indicator is computed from. The candidates are load-balancer access logs, in-process server metrics, the mobile client's own reporting, and synthetic probes. Which should be the primary source?
3 min answer sloslimeasurementavailability -
SLO Monitoring intermediate Multiple choice
A food-delivery platform in Zomato's mould pushes order-status events to restaurant tablets and to customers through a queue-backed webhook fleet. The complaints are that status arrives late rather than that it never arrives. Which indicator should the SLO be written on?
3 min answer slofreshnessasynchronouswebhooks -
Structured Logging intermediate
A platform moves from free-text logs to structured logs. What becomes possible, and what discipline must accompany it to avoid making things worse?
2 min answer structured-loggingschemacardinalityquerying -
Structured Logging intermediate
A ride-hailing platform of Grab's shape standardises on structured logs. After a release, the on-call dashboard that counts failed trips by city reads zero for three services and normal numbers for the rest. No errors are being reported anywhere. What broke, and what should have caught it?
2 min answer grabstructured loggingschema driftlog management -
Telemetry Cost intermediate Multiple choice
A design platform of the kind Canva runs exports 40 metrics per service. An engineer adds a `pod_name` label so a noisy pod can be identified. The service runs 600 pods and deploys twice a day, so pod names turn over completely every 12 hours. Retention is 30 days. Roughly how many distinct series does that one label create over the retention window?
2 min answer canvacardinalitytelemetry costmetrics -
Telemetry Cost intermediate
A platform's log volume has grown until log storage is one of its largest infrastructure costs. What should change, and what should not?
2 min answer elasticlogscostindexing