Field Service Technician II · Infrastructure · Reliability

Reliable infrastructure,built to be understood.

I’m Stephen McCruden, a Field Service Technician II working with mission-critical public-safety communications. I design, automate, and document systems with a focus on Kubernetes, GitOps, and operational clarity.

  • Kubernetes
  • Infrastructure as code
  • GitOps
  • Observability

Linux

Terraform

Ansible

Kubernetes

Flux

Prometheus

Engineering approach

Reliability is a design constraint.

01

Make it reproducible

A recovery plan is incomplete until the system can be reconstructed predictably from documented code.

02

Make behavior visible

Metrics, logs, and meaningful health signals turn guesswork into informed operational decisions.

03

Test the failure path

Backups, automation, and high availability only become trustworthy after deliberate failure testing.

From the notebook

Engineering notes and field lessons.

Browse all writing

See how I work

Systems are more credible when the decisions are visible.