← back to projects

Self-Service Platform @ project44

Golden-path tooling that lets product teams provision infrastructure and ship services without touching raw infra — across 50+ applications and 5 EKS clusters.

AWS EKSArgo CDTerraform HelmGitHub ActionsPrometheusGrafanaClaude Code

The problem: infrastructure as a bottleneck

When every team needs the platform team to hand-provision infrastructure and wire up deploys, the platform team becomes the bottleneck and every team reinvents the same wheel slightly differently. The fix isn't more tickets — it's paved roads: standardized, reusable building blocks teams can self-serve.

The golden path

I built the reusable layer that sits between raw AWS/Kubernetes and the product teams:

Product teams — a PR, not a ticket Golden-path building blocks Terraform modules Helm charts Argo CD ApplicationSets Actions templates 5 EKS clusters · multiple AWS accounts · 50+ apps Prometheus + Grafana · SLO dashboards · 99.95%

Observability I own

Availability doesn't come from hope — it comes from seeing problems before users do. I run the Prometheus + Grafana layer as dashboards-as-code with SLOs and tuned alerts, holding 99.95% service availability.

AI-assisted operations

I work AI-first. Claude Code runs as a daily agent in the infrastructure repos — but on a strict plan → generate → review-every-diff discipline. The judgment stays human; the toil gets automated.

RESULT

Incident investigation and troubleshooting time down ~30% — without letting an agent make an unreviewed change to IAM, state, or anything destructive.

outcomes

Where it landed

50+
apps on the platform
5
EKS clusters
99.95%
availability
~30%
faster incident triage