Skip to main content
Product

Infrastructure analysis — an X-ray of how you ship and run

I go through the way your code reaches production, the clusters it runs on, the bill it produces and what you can see when it breaks. You get a ranked list of what is genuinely urgent, what each option costs over three years, and a roadmap you can run yourself — or with me.

Report infrastructure analysis
Delivery path manual prod deploys
Kubernetes requests ≫ usage
Cloud cost untagged spend
Observability alerts tuned
Ranked by urgency · each option costed over 3 years

What you can expect

  • Read-only access — nothing changes while I look.
  • Findings ranked by urgency, not by what I could sell you next.
  • Scope agreed before I start, so there are no surprises.
  • The report is yours, whether or not you hire me afterwards.

Usually 1–2 weeks. Scope and price depend on the size of the setup, so they follow the first conversation.

What I look at

Four areas, the same four I work on as a consultant — read together, because the problems usually are.

01

The path to production

How a commit becomes a running release, and what happens when it has to be undone.

  • Manual steps and who can bypass them
  • Rollback — tested, or hoped for
  • Drift between environments
  • Secrets in pipelines
ArgoCDGitHub ActionsAzure DevOps
02

Kubernetes and the data on it

Whether the clusters are predictable — and whether the databases on them would survive a bad day.

  • Version and upgrade path
  • Requests vs real usage, autoscaling
  • RBAC and network policies
  • Postgres: HA, backups, point-in-time recovery
KubernetesHelmKarpenterCloudNativePG
03

Where the money goes

The bill read line by line, so every euro has an owner and a reason.

  • Spend by team and service
  • Idle and oversized resources
  • Commitments and spot
  • Lock-in premiums
KubeCostTerraform
04

What you can see when it breaks

Whether metrics, logs and traces answer the question you have during an incident — and who gets woken up.

  • Coverage vs what actually breaks
  • Alert noise and SLOs
  • Logs and traces joined up
  • On-call and runbooks
PrometheusLokiTempoGrafana

What you get

  • Written report

    Every finding with its evidence, ranked across security, reliability and cost.

  • Costed options

    Each fix with its effort and its cost over three years — including "leave it".

  • Roadmap

    Quick wins first, larger projects after, in an order your team can follow.

  • Walkthrough

    A session with your team to go through it, question it, and agree next steps.

If you want it built

The analysis stands on its own. When you want help with the fixes, this is the work it usually leads to.

  • GitOps with ArgoCD

    Every environment declared in Git and promoted by pull request.

    ArgoCD · Helm
  • Release automation

    One pipeline from commit to production: versioned, repeatable, easy to roll back.

    GitHub Actions · Azure DevOps
  • Infrastructure as code

    Terraform modules with Terragrunt per environment, and drift checks that tell you when reality moves.

    Terraform · Terragrunt
  • Postgres on Kubernetes

    CloudNativePG clusters with failover, backups to object storage and point-in-time recovery.

    CloudNativePG
  • Observability stack

    Metrics, logs and traces in one place, with dashboards and alerts worth waking up for.

    Prometheus · Loki · Tempo · Grafana

How it goes

  1. 01 Kickoff and access

    A 30-minute call, then read-only access to repos, clusters, the cloud bill and dashboards.

  2. 02 Analysis

    Usually 1–2 weeks. I read the pipelines, clusters, bill and telemetry — and ask your team what hurts.

  3. 03 Report and walkthrough

    Findings, costed options and a roadmap, walked through with your team.

  4. 04 Build (optional)

    Fix it with your team, with me, or both — the roadmap works either way.

Common questions

How long does an infrastructure analysis take?
Usually 1–2 weeks, depending on the size of the setup — from the kickoff call to the report walkthrough.
What access do you need?
Read-only access to your repositories, clusters, cloud bill and dashboards. Nothing changes while I look.
What do I get at the end?
A written report with every finding ranked by urgency, each fix costed over three years, a roadmap, and a walkthrough session with your team.
How much does it cost?
It depends on the size of the setup. Scope and price are agreed after a first 30-minute call, before any work starts.
Do I have to hire you afterwards?
No. The report is yours to keep. Your team can follow the roadmap alone, with me, or both.
Which platforms do you cover?
Kubernetes on AKS, EKS or GKE, Azure and AWS, and bare metal — with CI/CD in ArgoCD, GitHub Actions or Azure DevOps.

Want to know what is really going on?

Tell me what you run and what keeps going wrong. I'll come back with a scope for the analysis — no obligation.