Simple Life
9 days ago
Reach out directly about this role
Grade
Senior
Work Format
Remote
Employment
Full-time
English Level
B2 - Upper-Intermediate
By job title
AI writes it from your resume. You only have to send it.
Simple Life hiring Senior Site Reliability Engineer. Remote position. Simple is looking for a Senior SRE to join our Platform team responsible for the AWS infrastructure.
Simple is looking for a Senior SRE to join our Platform team responsible for the AWS infrastructure, the Kubernetes platform and the internal tooling that the rest of engineering relies on. Push the pace of innovation and build a future of a healthier world with us!
This is an operations-led position. You will be working day to day on AWS, our infrastructure-as-code, our CI/CD setup, observability, and the on-call rotation. A meaningful part of the work is automation: when we find ourselves doing the same thing twice, we usually invest in tooling rather than writing another runbook. Most of that tooling is written in Go.
To give a sense of the environment: infrastructure is defined with Terraform and Terramate, with Atlantis running plan and apply on pull requests. Workloads run on EKS with Karpenter, Cilium and Istio, deployed through ArgoCD. Observability is built on Grafana, Loki, Tempo, and Prometheus compatible metrics.
The most important quality for this role is how you handle problems that are not yet understood. Production incidents rarely present cleanly: logs can be incomplete, metrics can mislead, and the first plausible theory is often wrong. The right candidate stays focused under that kind of pressure, works through ambiguity in a structured way, and arrives at a real root cause rather than a convenient one. Strong investigation, debugging, and problem-solving instincts are essential.
We also expect candidates to learn quickly. The stack and the company both move, and you will regularly be the first person on the team to take on something new.