Senior Site Reliability Engineer

Garner

Remote regions

US

Salary range

$191,000–$226,000/yr

Benefits

Unlimited PTO

Similar Jobs

See all

What you’ll be part of:

  • Garner is on a mission to transform the U.S. healthcare system by applying proprietary clinical metrics to a dataset of 320M+ patients.
  • The company has helped over 2.5 million people access higher-quality care and saved $1B in healthcare costs.

About the role:

  • We are seeking a Senior Site Reliability Engineer to own the reliability, performance, and resilience of cloud infrastructure for Garner’s products and AI/ML workloads.
  • This role sits on the Platform Engineering team and involves defining SLOs, leading incident response, and driving automation.

What you will do:

  • Run the machine: own end-to-end reliability of cloud environments (AWS, Kubernetes) and define SLOs.
  • Lead incident response and drive deep-dive root cause analysis.
  • Automate away toil using AI tools to convert manual operational work into hands-free processes.

The ideal candidate has:

  • 4+ years of hands-on experience operating production cloud infrastructure at scale in an SRE, DevOps, or platform engineering role.
  • Deep expertise with Kubernetes and Terraform in a cloud-first environment (AWS preferred).
  • Strong software engineering fundamentals in Python or Go applied to infrastructure automation.

Garner

Garner partners with employers to redesign healthcare by using clinical metrics to identify top doctors and incentivize members to better care. The company has helped over 2.5 million people, saved $1B in costs, and doubled annually for five years, fostering a mission-driven, high-performance culture.

Apply for This Position