Platform Engineering
General material on internal platforms as products with users, adoption and lifecycles.
4 to work through
-
beginner Multiple choice
A platform injects a log-shipping sidecar with a 128 MiB limit into every pod. During an evening peak the sidecar's in-memory buffer grows faster than it can flush, the pod exceeds its limit, the kubelet kills the container and a product team's checkout service drops requests for eleven minutes. Whose reliability budget is spent and what does that imply about how the platform ships that sidecar?
3 min answer -
intermediate Multiple choice
What distinguishes platform engineering from an internal operations team?
2 min answer -
advanced
A platform team has built substantial capability and product teams report that it slows them down. What is the diagnostic, and what usually causes it?
2 min answer -
advanced
You are forming a platform team of four to serve twelve product teams. What do you build first?
1 min answer
3 terms in this topic
Pager Boundary
The written line in a platform contract that says which symptoms page the platform team and which page the consuming team, and how a single signal te…
conceptPlatform Engineering
Building internal products that reduce the cognitive load on delivery teams — with self-service as the property that separates it from a service desk.
practicePlatform User Research
Treating engineers as users whose actual behaviour is observed rather than assumed, which is what separates a platform from a set of shared tools.
Neighbouring topics
Internal Developer Platform
The assembled surface teams actually touch, and what belongs behind it.
Paved Road & Golden Path
A supported default route that is easier than the alternatives rather than mandatory.
Self-Service Provisioning
Teams getting infrastructure without a ticket, and the guardrails that make that safe.
Service Templates
Scaffolding new services with observability, CI and security already wired in.
Platform APIs
Treating the platform's own interfaces as contracts with consumers and compatibility rules.
Platform Tenancy
Isolating teams sharing a cluster, account or pipeline fleet, and where isolation must be hard.
Cluster Architecture
How many clusters, split by what, and the blast radius each split buys.
Service Mesh Operations
What a mesh genuinely solves, its failure modes, and the cost of running one.
Container Image Strategy
Base images, layer hygiene, rebuild cadence, and patching a fleet of images.
Developer Environments
Local, remote and ephemeral environments, and the fidelity each can honestly claim.
Inner Loop & Outer Loop
Where an engineer's time actually goes, and which loop a platform investment shortens.
Abstraction Level Choice
How much to hide, and the leak that turns a helpful abstraction into a trap.
Platform SLOs
Committing to reliability for internal consumers who cannot choose another provider.
Platform Adoption
Migrating existing teams onto a platform without a mandate, and reading the adoption curve.
Platform Funding
Central cost, showback, chargeback, and justifying a team that ships no customer feature.
Platform API Deprecation
Removing something dozens of internal teams depend on, on a timeline that holds.
Guardrails vs Gates
Preventing a class of mistake automatically versus stopping to ask a human.
Platform Telemetry
Instrumenting the platform itself: usage, friction, and where teams leave the paved road.
Platform Team Topologies
Stream-aligned, enabling, complicated-subsystem and platform teams, and their interactions.