Keep managed Kubernetes clusters healthy for customers who cannot afford downtime.
About the role
We operate Kubernetes for teams that would rather ship product than run control planes.
You will automate away toil, improve our observability stack and lead incident response.
What you will do
- Automate cluster provisioning and upgrades
- Own SLOs, alerting and incident retrospectives
- Improve the Prometheus and Grafana observability stack
What we are looking for
- Deep Kubernetes and Linux knowledge
- Terraform and CI/CD pipeline experience
- Calm, structured approach to incidents
Perks and benefits
- Fully remote
- Home office allowance
- Four-week sabbatical every three years
Skills
127 views