Quiz
2908 questions of the kind that actually get asked — in interviews, in architecture review boards, and by the person who has to run the thing at 3 AM. Every answer states the trade-off rather than the slogan, and says when the obvious choice is the wrong one.
All areas2908
Architecture Fundamentals91
Distributed Systems101
Data Architecture90
Cloud Architecture87
Networking86
API & Integration Architecture87
Reliability & Resilience99
Observability92
Performance & Capacity Engineering100
Security Architecture95
Cost Architecture & FinOps102
Business Architecture103
Architecture Communication102
Enterprise Architecture100
Legacy Modernization92
AI-Era Architecture96
Software Architecture & Engineering93
Architecture Patterns94
Architecture Decision-Making101
The Architect's Meta-Skills102
Delivery & Release Engineering103
Platform Engineering & Developer Experience102
Testing & Quality Architecture102
Data Platform Architecture98
Streaming & Real-Time Data103
Data Governance & Semantics91
Frontend & Experience Architecture101
Edge, Mobile & IoT98
Regulatory & Data Protection Architecture99
Assurance, Audit & Model Risk98
6 questions in Failure Thinking.
-
Failure Thinking intermediate
A major platform launch is three months away. How would you run a pre-mortem, and why bother?
2 min answer riskfacilitationmeta-skills -
Failure Thinking advanced
Between August and early September 2025 three separate infrastructure bugs degraded Claude's output quality - one affected 16% of Sonnet 4 requests in the worst hour of 31 August - while error rates and latency stayed normal. What class of failure is this, and what would you have had to build beforehand to notice it?
2 min answer anthropicsilent-failuredetectionquality-regression -
Failure Thinking advanced
Roblox's published account of its October 2021 outage - 73 hours from 28 to 31 October - says the difficulty of diagnosing two primarily unrelated problems buried deep in its Consul implementation was largely responsible for the length. Write-ups of that postmortem describe the remediation order: a node replaced on a degraded-hardware theory, then the whole cluster doubled from 64 to 128 cores with faster storage on a traffic theory, then a restore from an earlier healthy snapshot, then a return to 64-core hosts. What did that sequence assume, and what should change once the second remediation fails?
3 min answer robloxconsulincidenthypothesis -
Failure Thinking advanced
What does it mean to design by thinking about failure first, and what does that produce that requirement-driven design does not?
2 min answer credfailure-modesdesignpremortem -
Failure Thinking advanced
What does it mean to think in failure modes as a habit rather than as a checklist item, and which questions consistently produce findings?
3 min answer failure-modesdesign-reviewresiliencequestions -
Failure Thinking advanced
What questions characterise failure thinking, and which are most often skipped in design reviews?
2 min answer failure-thinkingreviewdependenciesdegradation