Source Job

$160,000–$208,000/yr
US Unlimited PTO

  • Design, build, and optimize reliable infrastructure for healthcare technology.
  • Improve scalability, reliability, and performance across distributed systems.
  • Collaborate with engineers and data professionals to shape modern infrastructure practices.

Python Go Kubernetes AWS Linux

20 jobs similar to Senior Site Reliability Engineer

Jobs ranked by similarity.

US

  • Build systems for declarative application and infrastructure lifecycle management, including CI/CD, Kubernetes, and service inventory.
  • Prioritize and troubleshoot infrastructure issues to minimize downtime and respond to alerts efficiently.
  • Contribute to setting the SRE team's direction and streamline automation of infrastructure processes.

Counterpart Health develops Counterpart Assistant, an AI-enabled primary care tool that supports physicians in chronic disease management. It is a subsidiary of Clover Health, with a remote-first culture and a focus on value-based care through technology.

$165,000–$216,000/yr
US Unlimited PTO

  • Develop internal tools and automate infrastructure using AWS, Kubernetes, and programming languages.
  • Research and design solutions to increase website robustness, availability, and cost efficiency.
  • Collaborate on documentation, code reviews, and rollout of new processes.

Angi powers the future of the home services industry, connecting homeowners with skilled pros. With 9 brands in 8 countries and employees worldwide, Angi has helped homeowners with over 300 million home projects.

$150,000–$195,000/yr
US Unlimited PTO

  • Design, secure, and scale modern cloud infrastructure to improve reliability and developer experiences.
  • Lead initiatives across infrastructure security, compliance, CI/CD, observability, and cloud operations.
  • Collaborate with engineering, IT, and business teams to solve complex security challenges and drive continuous improvement.

The company is a technology organization focused on building secure, scalable cloud infrastructure. It fosters a remote-first culture with an emphasis on innovation, ownership, and collaboration.

US Unlimited PTO 20w maternity 20w paternity

  • Design, build, and maintain highly available Kubernetes infrastructure at scale.
  • Lead design for components and features, and contribute to architecture decisions for container orchestration.
  • Mentor engineers on Kubernetes best practices and drive initiatives to improve system reliability.

Marqeta provides a card issuing platform for companies to issue cards, authorize transactions, and manage payment operations in real time. They are a publicly-traded company with a Flex First culture that values remote work and employee growth.

Poland

  • Design, write and deliver software to implement and support large web-scale, highly-performant, highly-available infrastructure on GCP/AWS.
  • Monitor infrastructure, respond to incidents, correct and improve systems to prevent incidents, and plan capacity.
  • Tune large-scale clusters for optimal performance and efficiency and support system deployments and product releases.

OpenX develops digital advertising marketplaces and technologies to optimize ad delivery for publishers and advertisers. The company operates a large-scale cloud infrastructure in Poland and values teamwork, customer centricity, and continuous learning.

$124,000–$170,500/yr
US

  • Perform operational deployments, implementations, and maintenance for production systems.
  • Implement and maintain monitoring, reporting, and alerting systems for Core Speech products.
  • Be part of an on-call rotation and work collaboratively to improve system performance and architecture.

Solventum is a new healthcare company with a long legacy of solving big challenges to improve lives and enable healthcare professionals to perform at their best. They are a large company that values empathy, insight, and clinical intelligence, collaborating with top minds in healthcare.

US Unlimited PTO

  • Keep user-facing services and production systems reliable, scalable, and efficient with automation and infrastructure-as-code.
  • Operate and troubleshoot production systems on Kubernetes, and contribute to observability with metrics, logs, and SLOs.
  • Participate in on-call, incident response, and post-incident reviews to drive improvements in automation and processes.

GitLab is the intelligent orchestration platform for DevSecOps, enabling organizations to increase developer productivity and improve operational efficiency. With more than 50 million registered users and over 50% of the Fortune 100 as customers, GitLab fosters a high-performance, all-remote culture driven by values and continuous knowledge exchange.

Canada

  • Design, implement, and maintain highly available and scalable infrastructure solutions.
  • Monitor system performance, identify bottlenecks, and resolve reliability issues proactively.
  • Automate infrastructure deployment, configuration management, and operational workflows.

The company is a technology firm that provides critical authorization solutions to organizations worldwide. It is a remote-first organization with a collaborative culture, offering equity opportunities and a focus on team building.

$140,400–$372,300/yr
US

  • Partner with engineering teams to improve reliability, scalability, and operational health of production systems.
  • Investigate and resolve complex production incidents, designing sustainable long-term solutions.
  • Design, build, and maintain automation tools and infrastructure to enhance developer productivity.

The company builds and maintains highly reliable, scalable production systems supporting millions of users worldwide. It fosters a remote-first culture that values innovation, collaboration, and engineering excellence.

$150,000–$200,000/yr
US Unlimited PTO

  • Design, build, and maintain scalable cloud infrastructure on AWS using Infrastructure as Code best practices.
  • Own and evolve the Atmos-based IaC framework and manage CI/CD pipelines with GitHub Actions.
  • Collaborate with engineers and scientists to support containerized and geospatial workloads.

Vibrant Planet develops a cloud-based AI platform to manage wildfire risk and modernize land management. They are a small, high-impact team backed by climate leaders, working on pressing climate challenges.

United States

  • Design and drive engineering-wide reliability programs to improve software quality and deployment confidence.
  • Build and enhance automated testing frameworks, CI/CD pipeline gates, and quality controls.
  • Develop scalable load testing solutions and production validation strategies for safe, high-confidence releases.

Jobgether uses an AI-powered matching process to ensure applications are reviewed quickly and fairly. They operate with a remote-first approach and collaborate with partner companies for hiring.

$152,000–$195,000/yr
US Unlimited PTO

  • Design, build, and scale Kubernetes infrastructure for secure, multi-tenant, high-availability applications.
  • Build and operate AI tooling infrastructure, including MCP servers and secure AI access.
  • Optimize CI/CD pipelines, implement progressive delivery, and advance Infrastructure as Code.

SecurityScorecard is the global leader in cybersecurity ratings, rating over 12 million companies across 64 countries. Headquartered in New York, it is recognized as a best workplace and funded by top investors.

$105,000–$115,000/yr
United States

  • Design, deploy, and manage highly available and secure cloud infrastructure across AWS and Azure.
  • Automate infrastructure provisioning using Terraform, CloudFormation, and ARM templates while implementing CI/CD pipelines.
  • Collaborate with development teams to build cloud-native solutions and ensure operational excellence through monitoring and improvements.

Jobgether is a job matching platform that uses AI-powered technology to connect candidates with employers. They facilitate remote hiring processes and support large-scale technology initiatives with a focus on efficiency and objectivity.

Global Unlimited PTO

  • Lead a high-impact infrastructure team, evolving internal platforms and CI/CD systems to support large-scale engineering operations.
  • Drive automation initiatives and AI-driven practices to reduce operational complexity and improve developer experience.
  • Define and execute strategies for scalable infrastructure, cloud environments, and platform engineering.

The partner company is a technology organization focused on building infrastructure platforms that enable engineering teams to deliver software faster. It is a remote-first company with a collaborative culture and a focus on innovation and scalability.

United States

  • Ensure reliability, scalability, and security of mission-critical cloud infrastructure and CI/CD environments.
  • Develop and implement automation solutions to streamline operational tasks and improve deployment efficiency.
  • Monitor system health and performance, proactively resolving incidents and contributing to continuous improvement.

Jobgether uses AI-powered matching to connect candidates with hiring companies. They are a platform that processes applications and shares top-fitting candidates with employers, operating remotely.

US Unlimited PTO

  • Architect and improve cloud foundations on Google Cloud Platform to support scalable, secure, and well-governed workloads.
  • Design and build platform capabilities across GCP, Kubernetes, CI/CD, GitOps, and developer tooling.
  • Mentor engineers and raise the technical bar through code review, architecture guidance, and direct implementation.

Wpromote is a digital marketing agency focused on performance marketing and technology. The company fosters a diverse, inclusive culture with a remote-friendly environment and office hubs in Los Angeles, Chicago, and New York.

$145,000–$175,000/yr
US 6w PTO

  • Leading design and implementation of robust, scalable, secure cloud-native solutions on AWS.
  • Developing and maintaining infrastructure-as-code for managing infrastructure across numerous Azure and AWS accounts.
  • Maintaining and optimizing CI/CD automation pipelines for rapid and reliable software deployments.

Element 84 is a woman-owned small business that develops geospatial data processing pipelines and builds software for public, private, and non-profit sectors. The company fosters a culture of curiosity and respect, supporting a large remote workforce with a flexible work schedule.

Brazil

  • Design and implement scalable cloud infrastructure solutions on AWS with automation and container orchestration.
  • Automate infrastructure provisioning using Terraform and manage Kubernetes clusters via Amazon EKS.
  • Develop CI/CD pipelines, monitor performance, apply security best practices, and document architecture for compliance.

CI&T helps large enterprises transform the potential of AI into real business impact with AI deployment, AI-native execution, and tech-integrated business solutions. With 30 years of experience and 8,000 CI&Ters across more than 25 countries, they accelerate innovation through a collaborative culture focused on Agentic SDLC, modernization, data, AI, martech, and business strategy.

US Unlimited PTO

  • Own the US-only production environment end-to-end, including infrastructure deployment, maintenance, scaling, and reliability.
  • Lead and grow the US-based DevOps team, design scalable AWS infrastructure, and build CI/CD pipelines for safe, fast shipping.
  • Partner with engineering on application error investigations, improve monitoring and alerting, and coordinate with the Tel Aviv team on shared platform standards.

Zafran de-risks 90% of critical vulnerabilities overnight across hybrid environments using existing security tools. Backed by Sequoia Capital and Cyberstarts, it is one of the fastest-growing companies in cybersecurity, scaling to meet demand from advanced organizations.

$170,000–$200,000/yr
US

  • Design and implement scalable, secure cloud infrastructure with a focus on FedRAMP compliance.
  • Build and optimize CI/CD pipelines and monitoring systems to enhance reliability and security.
  • Collaborate with engineering teams to automate infrastructure and improve operational excellence.