Platform engineering with an observability edge.
I build the automation, infrastructure-as-code and operating capabilities that let service teams ship, observe and own reliable systems across on-prem and GCP.
Miles Whittaker · Platform engineer · UK → Germany
I build the automation, operating paths and self-service capabilities that help teams run reliable systems — shaped by enterprise-scale observability and incident-management work.
About / 01
I’m Miles, a platform engineer at bet365 with a foundation in operations and infrastructure. Since moving into observability in 2024, I’ve led monitoring and incident-management transformation at enterprise scale — experience that now informs how I build platforms.
My work spans on-prem and GCP: replacing legacy monitoring, onboarding bespoke systems, and creating the automation and delivery paths that keep configuration consistent. I focus on more than tooling — clear ownership, practical training and self-service are what make platform change last.
I work across engineering teams and senior stakeholders to move response closer to the people who own each service: less central triage, less platform toil, and a clearer path from signal to action.
Experience / 02
Platform engineer building reliable, self-service foundations across hybrid environments — with enterprise-scale observability and incident-management transformation as a differentiating edge.
I build the automation, infrastructure-as-code and operating capabilities that let service teams ship, observe and own reliable systems across on-prem and GCP.
Primarily targeting Platform Engineering roles in Germany, with SRE and Forward Deployed Engineer roles also a strong fit where hands-on delivery, customer context and adoption matter. Berlin is a primary target.
Request the CV ↗Selected experience
Leading observability and incident-management transformation across on-prem and GCP: tens of thousands of hosts, 50+ teams, 800+ PagerDuty users and 3,000+ services.
Owned monitoring, incident triage, service continuity and change operations for mission-critical systems.
Terraform · Ansible · GCP · GitLab CI/CD · New Relic · PagerDuty · Telegraf · Grafana · Pyroscope · AWS · Kubernetes · Docker · Linux
Open to work
Based in the United Kingdom and actively exploring Germany-based Platform Engineering opportunities, alongside selected SRE and Forward Deployed Engineer roles. English is the authoritative site language for now.
Selected platform work / 03
Platform outcomes, delivered through observability: modernising monitoring at scale, making delivery repeatable, and shifting incident response to service owners.
Migrated observability across tens of thousands of hosts and hundreds of microservices. Automated rollout with Ansible and built bespoke integrations in Bash, PowerShell and Flex/cURL for systems standard agents could not cover.
Terraform and GitLab brought monitoring configuration under version control, making it repeatable and team-owned.
Replaced legacy Nagios ICMP checks with a highly available Telegraf-based service supporting hundreds of checks. Standardised rollout with Ansible to reduce configuration drift and simplify onboarding.
Built a GitOps pipeline to lint, test with telegraf --test, and deploy using Terraform, Ansible and GCP Cloud Deploy.
Led adoption across 50+ teams, 800+ users and 3,000+ services. Developed Terraform modules and Events API integrations, with service provisioning automated from CMDB and GKE discovery.
Routed alerts directly to service owners, reducing reliance on central triage. Improved alert quality with standards, grouping, hygiene and auto-pausing.
Interactive proof / 04
A deliberately generic, sanitised model of how platform changes and signals connect. Built as a teaching and onboarding aid; it does not depict my employer’s architecture or live systems.
Evidence lab / 05
A small, public-safe fixture showing how the same deployment can be investigated through profiles, traces and a change timeline.
Continuous profiling / Go
OpenTelemetry / Tempo
Release evidence / Git
r2026.09.12 baseline
Profile capturedsteady-state traffic · 60 s sampler2026.09.13 candidate
Template path changedprofile diff → trace check → promoteAll evidence here is synthetic and illustrative — not taken from a live service or employer telemetry. This demo shows an investigation workflow, not a production incident.
How I work / 06
The most useful systems are not just resilient. They explain themselves when the pressure is on.
Good work survives the person who made it. Interfaces, runbooks and architecture should leave the next decision easier than the last.
Observability is not a wall of charts. It is the shortest path from “something changed” to “we know why”.
Public surfaces get the polish. Private systems get the boundaries. A demo should be impressive without pretending to be live.
Start with a useful, measurable slice. Let the shape of the real problem earn the next layer of complexity.
Open channel / 07
For platform work, reliability engineering or a Germany-based opportunity, I’d like to hear what you’re building.
Start a conversation