Vantage platform rescue
Taking over undocumented infrastructure from a departed team and turning nightly incidents into a boring, reproducible platform.
The problem
Vantage's entire infrastructure had been built by one contractor who left without documentation. Deploys were manual, monitoring was a person watching a graph, and the team was averaging three incidents a week.
Nobody was confident they could rebuild it if a region went down.
What we did
We mapped what existed before changing anything, then codified it in Terraform piece by piece, verifying each step against the running system.
Deploys moved to a pipeline anyone could run. Monitoring moved from a person to alerts with runbooks attached to each one.
We finished by running a game day: deliberately destroying the staging environment and rebuilding it from code in front of the team.
The outcome
Incidents fell from roughly three a week to under one a month. Deploy time went from ninety minutes of manual work to six minutes in a pipeline.
The team can now rebuild the entire platform from an empty cloud account in under an hour.
We inherited infrastructure nobody understood after our contractor left. Seven weeks later we could rebuild the whole platform from code in under an hour. I sleep again.
Have something that needs to exist?
Send the brief. You get a scoped response with pricing, a delivery plan and the names of the people who would build it — within two working days.