Quiz
2667 questions of the kind that actually get asked — in interviews, in architecture review boards, and by the person who has to run the thing at 3 AM. Every answer states the trade-off rather than the slogan, and says when the obvious choice is the wrong one.
All areas2667
Architecture Fundamentals81
Distributed Systems101
Data Architecture90
Cloud Architecture87
Networking86
API & Integration Architecture78
Reliability & Resilience88
Observability81
Performance & Capacity Engineering90
Security Architecture95
Cost Architecture & FinOps92
Business Architecture93
Architecture Communication91
Enterprise Architecture91
Legacy Modernization92
AI-Era Architecture86
Software Architecture & Engineering84
Architecture Patterns84
Architecture Decision-Making91
The Architect's Meta-Skills92
Delivery & Release Engineering93
Platform Engineering & Developer Experience92
Testing & Quality Architecture90
Data Platform Architecture88
Streaming & Real-Time Data93
Data Governance & Semantics81
Frontend & Experience Architecture91
Edge, Mobile & IoT88
Regulatory & Data Protection Architecture90
Assurance, Audit & Model Risk88
90 questions in Performance & Capacity.
-
Performance & Capacity advanced
A distributed training job runs at 90% scaling efficiency on 512 GPUs. At 1,024 GPUs it drops to 60%. Walk me through where the time went, what you would measure, and what you would try.
3 min answer nvidiadistributed trainingcollectivesnccl -
Performance & Capacity intermediate
A search engine holds its index entirely in memory to guarantee predictable latency. What does that buy, what does it cost, and when is it the wrong choice?
2 min answer typesensemeilisearchin-memorypredictability -
Performance & Capacity advanced
Millions of customers attempt to buy a small number of items at a scheduled instant. Which performance constraint dominates, and why do conventional scaling techniques fail against it?
2 min answer flash-salecontentionhot-keyinventory -
Performance & Capacity intermediate
Peak trading day is six weeks away and expected to be four times normal traffic. What do you do in those six weeks?
3 min answer capacityload-testingpeakreadiness -
Performance & Capacity intermediate
Precompute every user's timeline at write time, or assemble it at read time? Explain why the answer for a social feed is neither.
2 min answer case-studytwitterfan-outpower-law -
Performance & Capacity intermediate
Your system handles 1,000 requests per second today. Marketing says a campaign will bring 10,000 next month. What breaks first, and how do you find out?
2 min answer capacitybottleneckload-testingscaling -
Profiling & Optimisation intermediate Multiple choice
A colleague spent a week optimising a function and the endpoint is 2% faster. What went wrong in the approach?
2 min answer optimisationamdahlprofilingmethod -
Profiling & Optimisation intermediate
A team has profiled a service and found the top three functions by CPU time. Why is optimising them often the wrong next step?
2 min answer profilingoptimisationamdahlmeasurement -
Profiling & Optimisation advanced
Discord's Read States service, written in Go, showed latency spikes every two minutes like clockwork. The team had written it carefully with very few allocations, and the spikes appeared regardless of load. Their published account from 2020 explains the cause and the rewrite that followed. What was happening, and what does it teach about periodic latency?
3 min answer discordgarbage collectiontail latencyruntime -
Queueing Theory intermediate
A capacity review shows services running at 85% CPU. Finance suggests raising it to 95% to save money. What is your response?
2 min answer capacitylatencycost -
Queueing Theory advanced Multiple choice
A platform runs its worker fleet at 85% average utilisation to control cost. Queue times are becoming unpredictable. What does queueing theory say is happening?
2 min answer browserstackqueueingutilisationlatency -
Queueing Theory advanced
An inference service's utilisation rises from 70% to 90% and latency more than triples. Why is the relationship non-linear, and what does that imply for capacity planning?
2 min answer queueing-theoryutilisationlatencyvariability