Source Job

US

  • Design and implement scalable cloud infrastructure to support growth.
  • Develop monitoring, alerting, and incident response for system reliability.
  • Automate deployment pipelines and ensure high availability and security.

AWS Terraform Docker Kubernetes Python

20 jobs similar to Site Reliability Engineer

Jobs ranked by similarity.

$105,000–$115,000/yr
US

  • Design, deploy, and manage scalable cloud infrastructure across AWS and Azure, including EC2, S3, RDS, and Kubernetes.
  • Automate cloud resources using Terraform, CloudFormation, and PowerShell, while implementing security controls and monitoring.
  • Collaborate with development teams to integrate CI/CD pipelines and support containerized workloads with Docker and Kubernetes.

Mind Computing supports the Department of Veterans Affairs by providing cloud engineering and DevOps solutions. The company fosters a collaborative culture with a focus on security, reliability, and innovation, though the team size is not specified.

Canada

  • Design, implement, and maintain highly available and scalable infrastructure solutions.
  • Monitor system performance, identify bottlenecks, and resolve reliability issues proactively.
  • Automate infrastructure deployment, configuration management, and operational workflows.

The company is a technology firm that provides critical authorization solutions to organizations worldwide. It is a remote-first organization with a collaborative culture, offering equity opportunities and a focus on team building.

Canada

  • Design, implement, maintain, and optimize highly available infrastructure supporting mission-critical applications and services.
  • Monitor production environments, analyze system performance, and proactively identify opportunities to improve stability, scalability, and operational efficiency.
  • Respond to technical escalations, troubleshoot infrastructure, networking, hardware, and software issues, and lead resolution of critical incidents.

Our partner is a technology company focused on high-availability platforms and mission-critical infrastructure. The team is collaborative and works with modern cloud technologies.

$160,000–$208,000/yr
US

  • Build systems for declarative application and infrastructure lifecycle management, including CI/CD, Kubernetes, and service inventory.
  • Prioritize and troubleshoot infrastructure issues to minimize downtime and respond to alerts efficiently.
  • Contribute to setting the SRE team's direction and streamline automation of infrastructure processes.

Counterpart Health develops Counterpart Assistant, an AI-enabled primary care tool that supports physicians in chronic disease management. It is a subsidiary of Clover Health, with a remote-first culture and a focus on value-based care through technology.

Global 6w PTO

  • Design, build, and maintain secure, scalable cloud infrastructure and CI/CD pipelines.
  • Implement monitoring, logging, and alerting systems to ensure operational excellence.
  • Collaborate across teams to improve system reliability, performance, and security.

The company operates an innovative B2B technology platform. It is a dynamic startup with a global team and a culture of ownership and autonomy.

$160,000–$208,000/yr
US Unlimited PTO

  • Design, build, and optimize reliable infrastructure for healthcare technology.
  • Improve scalability, reliability, and performance across distributed systems.
  • Collaborate with engineers and data professionals to shape modern infrastructure practices.

This company provides innovative healthcare technology solutions. It fosters a remote-first culture with a focus on engineering excellence and collaboration.

$105,000–$115,000/yr
United States

  • Design, deploy, and manage highly available and secure cloud infrastructure across AWS and Azure.
  • Automate infrastructure provisioning using Terraform, CloudFormation, and ARM templates while implementing CI/CD pipelines.
  • Collaborate with development teams to build cloud-native solutions and ensure operational excellence through monitoring and improvements.

Jobgether is a job matching platform that uses AI-powered technology to connect candidates with employers. They facilitate remote hiring processes and support large-scale technology initiatives with a focus on efficiency and objectivity.

Brazil

  • Act as a technical reference for SRE, DevOps, and cloud infrastructure, analyzing cloud environments in GCP/AWS for improvements and cost optimization.
  • Implement FinOps strategies, manage CI/CD pipelines, and maintain infrastructure as code using Terraform and Kubernetes.
  • Provide consultative support and communicate technical recommendations to engineering and business stakeholders.

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies. The company uses technology to streamline the application process and promote fair evaluation.

Brazil

  • Design and implement scalable cloud infrastructure solutions on AWS with automation and container orchestration.
  • Automate infrastructure provisioning using Terraform and manage Kubernetes clusters via Amazon EKS.
  • Develop CI/CD pipelines, monitor performance, apply security best practices, and document architecture for compliance.

CI&T helps large enterprises transform the potential of AI into real business impact with AI deployment, AI-native execution, and tech-integrated business solutions. With 30 years of experience and 8,000 CI&Ters across more than 25 countries, they accelerate innovation through a collaborative culture focused on Agentic SDLC, modernization, data, AI, martech, and business strategy.

US Unlimited PTO

  • Keep user-facing services and production systems reliable, scalable, and efficient with automation and infrastructure-as-code.
  • Operate and troubleshoot production systems on Kubernetes, and contribute to observability with metrics, logs, and SLOs.
  • Participate in on-call, incident response, and post-incident reviews to drive improvements in automation and processes.

GitLab is the intelligent orchestration platform for DevSecOps, enabling organizations to increase developer productivity and improve operational efficiency. With more than 50 million registered users and over 50% of the Fortune 100 as customers, GitLab fosters a high-performance, all-remote culture driven by values and continuous knowledge exchange.

$118,000–$151,000/yr
US 4w PTO

  • Build and improve platform services, including CI/CD pipelines and cloud infrastructure.
  • Collaborate with senior engineers to design scalable solutions and enhance developer experience.
  • Participate in incident response and retrospectives to drive continuous improvement.

Octopus Energy is a tech-powered energy company focused on renewable energy and customer experience. The company culture emphasizes ownership, collaboration, and making a tangible impact across teams.

US

  • Improve deployment reliability, reduce operational risk, and modernize AWS infrastructure toward Kubernetes.
  • Manage large-scale, high-availability distributed systems on AWS with Terraform and CI/CD pipelines.
  • Implement robust monitoring and observability, practicing SRE and security compliance best practices.

Peek is the operating system powering the experiences industry, helping museums, attractions, and tours increase revenues and deliver seamless guest experiences. Recognized by Forbes as a Best Startup Employer and by Built In as a Best Place to Work, we are a global remote-first team of Peeksters who obsess over customers and collaborate with purpose.

$100,000–$110,000/yr
US Unlimited PTO 14w maternity 14w paternity

  • Maintain uptime and security of AWS-hosted MERN applications and backend architectures.
  • Manage serverless functions and PySpark data pipelines, triaging incidents and automating workflows.
  • Lead blameless post-mortems and ensure HIPAA compliance across all healthcare data systems.

Cohere Health offers an AI-powered clinical intelligence platform connecting health plans and providers to improve care speed, cost, and quality. Named to the Inc. 5000 list and a top LinkedIn startup, the company is backed by leading investors and values empathy and growth.

Canada

  • Collaborate with Development and Architecture teams to build complex and highly available cloud environments.
  • Provide Level 3 technical support for internal teams, customers, and partners.
  • Design, implement, and maintain a secure and scalable infrastructure platform.

Smile Digital Health makes it easy for healthcare stakeholders to collect and exchange data with their FHIR-based data liberation platform. They were #19 on Deloitte's Technology Fast 50 Ranking for 2024 and foster a culture of respect, inclusion, and diversity.

Mexico

  • Design and promote modern DevOps practices to enhance collaboration across development, QA, and operations teams.
  • Deploy and manage Kubernetes clusters in cloud and on-premises environments for scalability and resilience.
  • Build and maintain CI/CD pipelines, automate infrastructure with IaC and CaC, and improve system observability.

Jobgether is a platform that uses AI-powered matching to connect candidates with job opportunities. It operates as a recruitment intermediary, partnering with companies to manage applications and next steps.

Global

  • Manage and optimize multi-cloud infrastructure (AWS required, GCP optional) with Kubernetes and CI/CD pipelines.
  • Improve observability through monitoring, logging, and alerting systems (e.g., Prometheus, Grafana, Coralogix).
  • Drive automation and Infrastructure as Code (IaC) using Terraform and Helm, and provide architectural guidance.

NIQ is the world's leading consumer intelligence company, delivering the most complete understanding of consumer buying behavior. In 2023, NIQ combined with GfK, bringing together two industry leaders with operations in 100+ markets and covering more than 90% of the world's population.

$150,000–$175,000/yr
US

  • Design and implement cloud infrastructure using AWS, Azure, and Terraform.
  • Manage Kubernetes clusters for container orchestration and deployment.
  • Collaborate with cross-functional teams to ensure scalability and reliability.

BSC Analytics is a leader in advanced data analytics for highly regulated enterprises, providing technical strategy and teams of exclusively senior talent. The company fosters a culture of senior expertise, tackling the toughest data challenges.

$150,000–$195,000/yr
US Unlimited PTO

  • Design, secure, and scale modern cloud infrastructure to improve reliability and developer experiences.
  • Lead initiatives across infrastructure security, compliance, CI/CD, observability, and cloud operations.
  • Collaborate with engineering, IT, and business teams to solve complex security challenges and drive continuous improvement.

The company is a technology organization focused on building secure, scalable cloud infrastructure. It fosters a remote-first culture with an emphasis on innovation, ownership, and collaboration.

$46,000–$150,000/yr
US

  • Design and maintain CI/CD pipelines for secure software delivery.
  • Manage containerized applications using Docker and Kubernetes.
  • Automate cloud infrastructure with Infrastructure as Code tools such as Terraform.

Clarity Innovations is a trusted national security partner providing innovative solutions for the Intelligence Community and Department of Defense. They are a people-focused company committed to being a destination employer for top talent.

US Unlimited PTO

  • Architect and improve cloud foundations on Google Cloud Platform to support scalable, secure, and well-governed workloads.
  • Design and build platform capabilities across GCP, Kubernetes, CI/CD, GitOps, and developer tooling.
  • Mentor engineers and raise the technical bar through code review, architecture guidance, and direct implementation.

Wpromote is a digital marketing agency focused on performance marketing and technology. The company fosters a diverse, inclusive culture with a remote-friendly environment and office hubs in Los Angeles, Chicago, and New York.