Source Job

Global Unlimited PTO

  • Own infrastructure end to end, including Kubernetes clusters, Docker builds, and DigitalOcean services.
  • Move production deploys toward safer, more frequent releases and manage capacity, autoscaling, and load balancing.
  • Operate PostgreSQL, Redis, and NATS under real traffic, keep Cloudflare tight, and document systems for the whole team.

Kubernetes Terraform PostgreSQL Redis Cloudflare

20 jobs similar to DevOps Engineer

Jobs ranked by similarity.

US

  • Implement and manage the infrastructure stack to enable the engineering team to ship quickly and effectively.
  • Proactively identify and eliminate bottlenecks in the devops process to ensure optimal developer velocity.
  • Maintain Tempo chain reliability, validator infrastructure, and explorer reliability.

Tempo is a layer-1 blockchain purpose-built for stablecoins and real-world payments, born from Stripe and Paradigm. They are a team of crypto-optimists building infrastructure for onchain payments.

Brazil 4w PTO

  • Own critical infrastructure across compute, networking, CI/CD, Kubernetes, and observability.
  • Manage Kubernetes environments and infrastructure-as-code with Terraform, improving developer experience and reducing operational friction.
  • Lead production incident response, influence architecture, and integrate AI-powered tools to boost engineering efficiency.

Jobgether is an AI-powered recruitment platform that connects candidates with global hiring companies. This role is with a partner company, a globally distributed technology organization offering a collaborative, informal culture and long-term opportunities.

Global

  • Build and deploy DevOps tooling using infrastructure-as-code and container orchestration to power developer portals.
  • Own the full lifecycle of DevRel infrastructure including Kubernetes, GitOps, and high-availability setups.
  • Implement CI/CD pipelines, monitoring, and automation for documentation and code samples.

Vultr provides high-performance cloud infrastructure including Cloud Compute, Cloud GPU, Bare Metal, and Cloud Storage for enterprises and AI innovators worldwide. It is the world's largest privately-held cloud infrastructure company, trusted by hundreds of thousands of customers across 185 countries, with a culture focused on employee care and inclusion.

$59,400–$65,880/yr
Europe

  • Lead the design, implementation, and ongoing improvement of reliable, scalable, and secure production platforms and services.
  • Work closely with cross-functional teams to build and maintain resilient infrastructure and deployment patterns.
  • Provide technical leadership and mentorship, promoting strong engineering standards and operational best practices.

Cision is a global leader in PR, marketing and social media management technology and intelligence, helping brands connect with customers and stakeholders. They have offices in 24 countries, a network of over 1.1 billion influencers, and a culture that champions diversity, equity, and inclusion.

US Unlimited PTO

  • Build and operate the Kubernetes platform supporting AI test and evaluation frameworks.
  • Design infrastructure-as-code, GitOps workflows, and automated deployment pipelines.
  • Own platform reliability, observability, capacity planning, and operational readiness.

OpenTeams helps enterprises and governments build AI they control, govern, and evolve themselves. Founded by the creator of NumPy and SciPy, the company is built by people with deep roots across the open-source ecosystem and maintains a remote-first culture.

Europe

  • Own and evolve production infrastructure, leading the migration from Docker Swarm to Kubernetes on premises.
  • Drive observability, enforce IaC practices, and ensure CI/CD reliability across ~50 services.
  • Participate in on-call rotation, resolve incidents, and build platform tooling to reduce infrastructure toil.

Webshare is an enterprise-grade proxy platform providing access to over 80 million global IPs across 195 countries. With 99.97% uptime, it serves tens of thousands of businesses, and the team emphasizes mentorship, knowledge-sharing, and team events.

$220,000–$292,000/yr
US Unlimited PTO

  • Own the platform including GCP, Kubernetes, Temporal, GPU fleet, and deploy/rollback machinery.
  • Contribute to AI enablement substrate: GPU capacity, training/inference pipelines, and cost optimization.
  • Strengthen team practices through tooling, standards, tests, observability, and release processes.

Descript is building a simple, intuitive, fully-powered editing tool for video and audio — an editing tool built for the age of AI. They are a team of 150 backed by top investors like OpenAI and Andreessen Horowitz, with a culture that values collaboration and serendipitous discovery.

Canada USA Unlimited PTO

  • Own the observability, logging and alerting for Kubernetes clusters and critical workloads.
  • Build and maintain automation for lifecycle management of Kubernetes clusters.
  • Identify and root-fix reliability bottlenecks before they become incidents.

Wrapbook is an AI platform for production finance, built for feature films and TV, trusted by Netflix and Paramount. Backed by top investors, our team of over 350 employees uses AI to transform how finance teams work.

US Unlimited PTO

  • Own the infrastructure layer for AI workloads including inference serving, Kubernetes, and agent-sandboxing platforms.
  • Manage the serving tier for open-weight models, Kubernetes operators, and stateful data planes.
  • Oversee the sandbox runtime, control-plane services, and observability tooling.

AZX accelerates positive impact in critical industries through AI transformation, specializing in physics-informed ML and enterprise AI solutions for climate and sustainability. Founded in 2024, the company is a profitable public benefit corporation with a growing team working with category leaders in real estate, energy, logistics, and utilities.

$126,000–$174,000/yr
US

  • Define DevOps strategy and lead infrastructure architecture across multi-environment, multi-region cloud systems.
  • Architect and own scalable Kubernetes platforms, infrastructure as code, and DevSecOps implementation.
  • Drive platform reliability, performance SLAs, cost optimization, and lead complex migrations and AI/ML platform infrastructure.

Robots & Pencils is an applied AI engineering firm that designs and ships AI co-workers for enterprise operations. Founded in 2009, the company has delivery centers across Canada, the US, Eastern Europe, and Latin America, with teams averaging over 15 years of experience.

Europe

  • Operate and improve Linux infrastructure and Kubernetes clusters across bare-metal, virtualized, and on-premise environments.
  • Design and maintain complex networking architectures and automation using Ansible, Bash, Python, and GitOps.
  • Lead incident response, define SLOs, and build observability platforms with Prometheus, Grafana, and ELK.

Jobgether is a platform that connects job seekers with opportunities through an AI-powered matching process. The company fosters a remote-first culture and emphasizes autonomy and ownership for engineers.

US

  • Define the strategy, roadmap, and feature priorities for k0rdent AI Kubernetes services, empowering Neocloud operators to launch managed Kubernetes on their own GPU infrastructure.
  • Translate requirements from NeoClouds, GPU clouds, telcos, and enterprise platform teams into clear product direction and partner with engineering to ship secure, scalable cluster lifecycle capabilities.
  • Manage the Kubernetes backlog, define positioning and competitive differentiation, and create field-facing assets to support strategic accounts.

Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI. They have a world-class, distributed team committed to openness and technical excellence.

$135,000–$150,000/yr
US

  • Operate, scale, and troubleshoot Bitsight's SaaS cloud infrastructure with focus on reliability, efficiency, and security.
  • Tackle complex system-level designs and proactively anticipate performance and scalability issues.
  • Pioneer self-optimizing infrastructure systems using AI, ensuring manual and staging validation before production deployment.

Bitsight is a cyber risk management leader transforming how companies manage exposure, performance, and risk. Over 3,500 customers and 600 teammates work across Boston, Raleigh, New York, Lisbon, Singapore, and remote locations.

$139,200–$235,200/yr
Canada United States Unlimited PTO

  • Design, build, and operate GitLab Orbit backend services, primarily in Rust, within a distributed, cloud-native environment.
  • Improve deployment, monitoring, and operations using Kubernetes, Helm, Terraform, and cloud services from AWS or GCP.
  • Automate operational work, strengthen observability, and manage production issues to reduce single points of failure.

GitLab is the intelligent orchestration platform for DevSecOps, helping organizations increase developer productivity, improve operational efficiency, and accelerate digital transformation. Trusted by more than 50 million registered users and over 50% of the Fortune 100, GitLab fosters a high-performance culture driven by shared values and continuous knowledge exchange.

Global

  • Design and implement scalable Kubernetes infrastructure for high-throughput event processing.
  • Build cloud-agnostic environments and implement GitOps workflows using Terraform.
  • Manage databases, monitoring, and security in production Kubernetes environments.

Miratech is a global IT services and consulting company that helps visionaries change the world. The company retains nearly 1000 full-time professionals and has a culture of relentless performance with a 99% project success rate.

$115,000–$130,000/yr
Americas

  • Design and maintain data pipelines that ingest and serve data across Techstars systems.
  • Own Kubernetes-based infrastructure, deployment tooling, and cluster health.
  • Maintain PostgreSQL database layer and support Salesforce data flows.

Techstars is a global investment firm that backs ambitious founders with capital, mentorship, and the connections they need to grow. The company has operated since 2006 and offers a collaborative, inclusive environment.

$150,000–$165,000/yr
US Unlimited PTO

  • Design and manage high-availability platforms using Kubernetes, Terraform, and Ansible with native-AI capabilities.
  • Develop and operate the observability stack: Grafana, Mimir, Loki, Tempo, and Prometheus on Kubernetes via GitLab CI/CD.
  • Build automation scripts in Python, maintain GitOps pipelines, and mentor mid-level engineers.

Flexential builds and operates critical IT platforms including observability, DevOps, and ITSM technologies. The company fosters a collaborative engineering culture and values diversity.

EU

  • Own and operate production infrastructure across Kubernetes, Linux, networking, and virtualization.
  • Lead incident response and implement observability to improve availability and performance.
  • Define SLOs and automate infrastructure with Ansible, Bash, Python, and GitOps.

Jobgether is an AI-powered job matching platform that connects candidates with hiring companies through objective, data-driven processes. They foster a collaborative, international, and fully remote work environment, emphasizing autonomy and ownership for their small to mid-sized team.

  • Own the reliability, performance, and scalability of Runlayer's infrastructure across AWS and GCP.
  • Manage Kubernetes clusters, database reliability, and CI/CD pipelines for rapid deployments.
  • Lead incident response and partner with product engineers to design resilient systems for enterprise customers.

Runlayer builds a unified platform for MCPs, Skills, and AI Agents, providing enterprises with security, governance, and observability to deploy AI safely and at scale. Founded by engineers who built AI Actions for OpenAI and Zapier Agents, the team has raised $42M from Felicis and Khosla Ventures, serving companies like Gusto, Instacart, and Opendoor.

Global

  • Contribute to platform and harness engineering, including CI/CD and developer tooling.
  • Build systems to reduce toil and maintain production infrastructure under conversational AI traffic.
  • Participate in on-call rotation and incident management to ensure platform uptime.

Replicant builds an AI-powered customer service platform that helps contact centers resolve requests and improve agent performance. The company is distributed, with a focus on ownership and collaboration, and serves Fortune 500 companies.