Remote Devops Jobs · Kubernetes

Job listings

$140,000–$170,000/yr

  • Deploy and operate Blitzy's self-hosted platform within a customer-controlled, secure cloud environment.
  • Own the Kubernetes-based deployment, releases, upgrades, capacity planning, and performance benchmarking.
  • Serve as the on-account technical presence, partnering with customer infrastructure and security teams.

We are an AI software development platform that autonomously builds custom software for enterprises. Backed by tier 1 investors and led by two co-founders, we are one of the fastest-growing U.S. companies with a culture of speed and customer focus.

  • You'll operate production day-to-day, including oncall, incident response, and postmortems.
  • You'll own reliability practice by defining SLIs/SLOs and error budgets.
  • You'll ship infrastructure through code in a GitOps workflow for cloud and Kubernetes.

Alpaca is a global leader in agent-first brokerage infrastructure for stocks, ETFs, options, crypto, and more. With a team of 400+ globally distributed members, we foster a culture of curiosity, empathy, and accountability.

$140,000–$165,000/yr
Global Unlimited PTO

  • Design, build, and optimize cloud infrastructure (AWS/Kubernetes/EKS) and CI/CD pipelines across multiple teams.
  • Troubleshoot and resolve production incidents of varying scope, ensuring reliability and performance.
  • Drive infrastructure projects end-to-end, mentor engineers, and establish standards that improve developer productivity.

Pacvue is a leading Commerce Media OS powering over $12B in advertising spend across 100+ global retail media networks. It enables over 70,000 brands and agencies with an inclusive global community that fosters innovation and career growth.

  • Engage directly with customers to resolve complex technical challenges involving Kubernetes GPU clusters.
  • Act as a customer-facing SRE to ensure Kubernetes clusters remain healthy and stable.
  • Become a product expert in GPU Cluster service, serving as the last line of technical defense before escalation.

Together AI is a research-driven artificial intelligence company focused on open and transparent AI systems. The team has contributed to leading open-source research and aims to build the next generation AI infrastructure.

$115,000–$175,000/yr
Global Unlimited PTO

  • Own and operate the AWS and Kubernetes platform, taking responsibility for reliability, cost, and performance.
  • Lead GitOps deployments and manage CI/CD end-to-end with GitLab pipelines.
  • Establish platform observability and maintain security baselines.

Veeva Systems is a mission-driven organization and pioneer in industry cloud, helping life sciences companies bring therapies to patients faster. It is one of the fastest-growing SaaS companies in history, surpassing $3B in revenue in its last fiscal year, with a culture focused on values like Do the Right Thing, Customer Success, Employee Success, and Speed.

$185,000–$230,000/yr
US Unlimited PTO

  • Architect, automate, and maintain Kubernetes deployments across commercial, CUI, and classified DOD networks.
  • Debug complex containerized environments where connectivity is limited.
  • Interface directly with customers and cross-functional teams to define and deploy reliable systems.

Striveworks delivers trusted AI systems for government and enterprise, monitoring performance and managing drift across hundreds of deployed models. The company values trust, respect, and ownership, with a collaborative culture focused on mission-critical reliability.

  • Own the infrastructure end-to-end for ScaleOps' self-hosted and SaaS platforms.
  • Manage cloud infrastructure across AWS, GCP, and Azure, including networking, security, and compute.
  • Collaborate with customers and internal teams to ensure rapid feature delivery without compromising reliability.

ScaleOps is redefining autonomous cloud and AI infrastructure, freeing DevOps from manual resource management. Backed by $210M+ in funding, they are trusted by leading enterprises and Fortune 100 companies, with a fast-paced, innovative culture.

$75,600–$124,200/yr

  • Design and maintain highly available, scalable systems to ensure exceptional customer experiences.
  • Drive automation and eliminate operational toil through self-service tooling and process improvements.
  • Lead incident response and mentor engineers to improve reliability practices.

Redzone provides a connected workforce solution for manufacturers to improve plant efficiency and worker productivity. The company is part of QAD Inc. and fosters a collaborative, customer-focused culture with a strong technology team.

  • Own infrastructure as code using Terraform and Terragrunt for scalable, reliable cloud infrastructure.
  • Design and optimize CI/CD pipelines in GitLab and manage Kubernetes workloads for microservices.
  • Drive security, reliability, and collaborate with development teams to ensure compliance and performance.

Deutsche Telekom IT Solutions Slovakia provides innovative information and communication technology services. The company has grown to over 3900 employees and is the second largest employer in eastern Slovakia, promoting a culture of continuous improvement and work-life balance.

  • Deploy, integrate, and operate high-performance storage for GPU-accelerated compute and AI platforms.
  • Own the storage layer where Kubernetes meets bare metal, tuning NFS data paths for high-throughput workloads.
  • Automate storage provisioning with infrastructure-as-code and GitOps, ensuring observability and reliability.

Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI and data-intensive applications. With deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams across hybrid, edge, and sovereign environments, fostering a culture of open-source innovation and collaboration among passionate, talented colleagues.