Site Reliability Engineer

Skyline CloudPunePosted last week

RemoteFull-time3 - 5 years₹20 - 30 LPA

Keep managed Kubernetes clusters healthy for customers who cannot afford downtime.

About the role

We operate Kubernetes for teams that would rather ship product than run control planes. You will automate away toil, improve our observability stack and lead incident response.

What you will do

  • Automate cluster provisioning and upgrades
  • Own SLOs, alerting and incident retrospectives
  • Improve the Prometheus and Grafana observability stack

What we are looking for

  • Deep Kubernetes and Linux knowledge
  • Terraform and CI/CD pipeline experience
  • Calm, structured approach to incidents

Perks and benefits

  • Fully remote
  • Home office allowance
  • Four-week sabbatical every three years

Skills

127 views

Similar openings

Featured

Own and scale the settlement APIs that move money for thousands of merchants every single day.

Bangalore
1 - 3 years
₹14 - 22 LPA
  • Java
  • Spring Boot
  • MySQL
  • Kafka
  • +1
HybridPosted 2 days ago

Define the internal platform that every product team at Cortex builds on.

Remote
5+ years
₹35 - 50 LPA
  • Kubernetes
  • Go
  • Platform Engineering
  • Observability
RemotePosted 4 days ago