Observability, SRE & AI Leader
Hands-on leadership. Shared know-how.
Building observability platforms on Kubernetes, managing infrastructure as code and engineering the tools around AI. Helping other engineers work through technical decisions, and turning incident findings into changes to code, checks and runbooks.
Take a look aroundSelected work
Under the surface.
Observability beyond
the dashboard.
Telemetry collection, continuous profiling and production investigation. From Kubernetes resource sizing to whether a health check tells the truth.
Beyond the
prompt window.
Go MCP tooling, TypeScript agent extensions and a shared LLM proxy. Session continuity, useful integrations and the limits around them.
Kubernetes & IaC.
Through to production.
Git-managed Kubernetes configuration, Helm delivery and Terraform state migration. Reviewed changes, explicit promotion and recovery paths.
Make it useful
to the next person.
Incident-management requirements, routing standards and vendor collaboration, alongside mentoring, implementation reviews and shared operational guidance.
Not everything needs a runbook
Off the clock.
Still curious.
There is life outside a terminal. Theatre, cinema and gaming, for a start.
Some side quests do involve a server.
Self-hosted services for family and friends give the tinkering a practical home. A separate Kubernetes and GitOps lab leaves room to explore observability, try things out and get things wrong without interrupting anyone's evening.