Source Job

UK

  • Manage and optimize Kubernetes clusters in GKE through Terraform.
  • Design and implement golden paths and automation strategies that empower developers to self-serve.
  • Serve as the technical point-of-contact for GCP and Kubernetes queries, supporting compliance and internal teams.

Kubernetes Terraform Google Cloud Platform Python

20 jobs similar to Cloud Platform Engineer

Jobs ranked by similarity.

UK Ireland Estonia Netherlands Sweden Israel Eastern Europe Portugal Unlimited PTO

  • Develop and evolve foundational software and services enabling product and development teams.\n- Architect, design, and implement Infrastructure as Code using Terraform.\n- Deploy, manage, and optimize Kubernetes clusters on GCP (GKE) and AWS (EKS).

DoiT is a global technology company that helps cloud-driven organizations leverage cloud for business growth and innovation through data, technology, and human expertise. They work with over 4,000 customers worldwide and foster a remote-first, entrepreneurial culture.

France

  • Evolve an Internal Developer Platform enabling development teams to deploy and operate applications securely with high self-service.
  • Play a crucial role in infrastructure architecture in a multi-cloud environment with a focus on GCP.
  • Build and maintain the platform used by over 50 internal clients, then support their concrete use by teams.

Lifen believes medical data can transform healthcare by reducing administrative burden, improving care coordination, and accelerating scientific discovery. Since 2015, the company has connected 800 hospitals and 150,000 healthcare professionals, with over 150 employees working remotely and from offices to unlock the potential of health data.

Europe Unlimited PTO

  • Design and build scalable, reliable cloud infrastructure on GCP and AWS.
  • Manage Kubernetes environments and infrastructure as code with Terraform.
  • Drive CI/CD automation, platform reliability, and developer self-service.

The company builds and operates scalable cloud infrastructure and internal developer platforms. It is a globally distributed, fully remote engineering team with a collaborative and inclusive culture.

$175,000–$185,000/yr
US

  • Consolidate Terraform and establish conventions for state management, modules, and CI checks.
  • Improve monitoring, observability, and automation in Datadog and Cloud Monitoring.
  • Right-size workloads, evaluate Kubernetes architecture, and retire legacy tooling.

Vida is a virtual, personalized obesity care provider that combines evidence-based treatment with advanced technology to help patients improve their health. Trusted by Fortune 100 companies and growing for years, Vida takes a whole-person approach to care and celebrates diversity across its team.

Canada USA Unlimited PTO

  • Own the observability, logging and alerting for Kubernetes clusters and critical workloads.
  • Build and maintain automation for lifecycle management of Kubernetes clusters.
  • Identify and root-fix reliability bottlenecks before they become incidents.

Wrapbook is an AI platform for production finance, built for feature films and TV, trusted by Netflix and Paramount. Backed by top investors, our team of over 350 employees uses AI to transform how finance teams work.

$185,000–$200,000/yr
US Unlimited PTO 12w maternity 12w paternity

  • Coordinate with technical and non-technical staff across departments, including workflow automation that bridges infrastructure and business processes.
  • Design, implement, and maintain scalable, secure, and highly available cloud infrastructure in GCP.
  • Maintain incident response process and tooling, and build automation that reduces toil and enables self-healing infrastructure.

Branch empowers workers with financial freedom by helping companies accelerate payments and providing accessible, free financial services. It is a remote-first, award-winning FinTech with employees across the U.S., fostering a culture of transparency, accountability, and trust.

$190,000–$230,000/yr
US

  • Design and implement scalable cloud infrastructure using Kubernetes, Pub/Sub, and distributed systems technologies.
  • Collaborate with our AI team to optimize data pipelines and integrate AI to remove performance bottlenecks.
  • Drive platform reliability initiatives including alerting, health checking, and incident management.

Syllo is building a unified litigation platform that helps lawyers and paralegals use AI throughout the litigation life cycle. We are a quickly expanding company with enterprise customers including major law firms and corporations.

Canada

  • Support production systems across Azure, GCP, and datacenters to meet SLA targets.
  • Automate with Terraform, GitHub Actions, ArgoCD, and Python/Bash/PowerShell.
  • Participate in on-call rotation and collaborate to resolve incidents and improve reliability.

Kinaxis is a global leader in modern supply chain orchestration, with an AI-infused platform that provides end-to-end visibility. Starting as a team of three in 1984, Kinaxis now has over 2,000 employees globally and is known for its strong culture and technology.

India

  • Architect and scale multi-region microservices, APIs, and authentication infrastructure on AWS/GCP.
  • Lead SLOs, observability, incident management, and disaster recovery automation to maintain 99.99% availability.
  • Manage Kubernetes clusters and Terraform IaC while eliminating toil with Python/Go tooling.

JumpCloud is an AI-powered unified IT management platform that secures the modern workforce through identity, device, and access management. The company is remote-first with teams in 15+ countries and values building connections, thinking big, and continuous improvement.

$145,000–$145,000/yr
US

  • Design, deploy, and maintain GCP infrastructure, including Compute Engine, GKE, Cloud Storage, IAM, and Cloud Interconnect, following well-architected principles.
  • Automate provisioning using Terraform, implement IAM best practices, and partner on security audits and hardening.
  • Monitor performance, availability, and cost, and support cross-cloud and on-premises networking as needed.

Dragos is the global leader in xOT cybersecurity, protecting critical infrastructure systems that deliver water, power, and healthcare. The remote-first team spans North America, Europe, the Middle East, and APAC, built on authenticity, transparency, and trust.

Global

  • Develop, secure and maintain cloud infrastructure for large-scale training, inference and evaluation platforms.
  • Design and implement CI/CD pipelines, reusable infrastructure templates and automation for rapid environment creation.
  • Collaborate with engineering and research teams to streamline releases and improve developer workflows.

We build intelligent systems capable of learning, remembering and reasoning over time. We are an early-stage startup with ambitious goals and a highly collaborative, fast-moving environment.

$180,000–$220,000/yr
US

  • Design, implement, and maintain reliable, scalable, and secure infrastructure to support applications and automation systems.
  • Automate infrastructure provisioning, configuration management, and deployment pipelines using tools like Terraform and ArgoCD.
  • Implement observability solutions and enforce security best practices to ensure uptime and system performance.

Bright Machines is a next-generation, AI-enabled manufacturer focused on data center infrastructure production, using proprietary AI-based robotics and software to assemble hardware products for hyperscalers and OEMs. The company is headquartered in San Francisco, California, with an integration center in Guadalajara, Mexico, and has been recognized by Forbes' AI 50 and other leading organizations.

India

  • Design, build, and ship production services, APIs, and user-facing interfaces.
  • Build and operate production AI systems including RAG, fine-tuning, and inference optimization.
  • Architect AWS/GCP environments with Kubernetes and Terraform and control cloud/AI costs.

Motive empowers people who run physical operations with tools to make their work safer, more productive, and more profitable. Serving nearly 100,000 customers across industries, the company values a diverse and inclusive workplace.

Europe

  • Lead end-to-end technical engagements: Partner directly with engineering teams to diagnose, unblock, and resolve complex infrastructure challenges.
  • Execute critical migrations: Develop reference implementations, tooling, and guidance to transition teams off deprecated systems seamlessly.
  • Accelerate platform adoption: Act as primary technical contact for new teams onboarding to Planet's core infrastructure.

Planet designs, builds, and operates the largest constellation of imaging satellites in history, delivering unprecedented dataset via a cloud-based platform for commercial, environmental, and humanitarian sectors. A global company with offices in the US, Europe, and Slovenia, Planet values a people-centric culture and community.

US

  • Own ScaleOps' infrastructure end-to-end, including self-hosted product, SaaS platform, and AI infrastructure.
  • Manage cloud infrastructure across AWS, GCP, and Azure, covering networking, security, SSO, and compute.
  • Collaborate with customers and internal teams to ensure reliable feature delivery and eliminate operational toil.

ScaleOps is redefining autonomous cloud and AI infrastructure, freeing DevOps and platform engineers from manual resource management. We are the category leader backed by over $210M in funding, trusted by leading enterprises including Adobe, Coinbase, and Fortune 100 companies.

$135,000–$150,000/yr
US

  • Operate, scale, and troubleshoot Bitsight's SaaS cloud infrastructure with focus on reliability, efficiency, and security.
  • Tackle complex system-level designs and proactively anticipate performance and scalability issues.
  • Pioneer self-optimizing infrastructure systems using AI, ensuring manual and staging validation before production deployment.

Bitsight is a cyber risk management leader transforming how companies manage exposure, performance, and risk. Over 3,500 customers and 600 teammates work across Boston, Raleigh, New York, Lisbon, Singapore, and remote locations.

US

  • Own core platform infrastructure including Terraform migration, CI/CD pipelines, and environment provisioning as part of a growing team.
  • Build and maintain tooling and abstractions that let product engineers ship and run code without solving infrastructure problems from scratch.
  • Advance observability foundation, standardize monitoring and alerting, and support GCP infrastructure scaling.

Astra builds mission-critical infrastructure for moving money at scale, processing billions in annual transaction volume with 99.9%+ uptime. We are a remote-first company hiring within the U.S., with a small team focused on thoughtful collaboration and clarity.

$135,000–$150,000/yr
US

  • Build, deploy, and maintain secure GCP environments across Compute Engine, GKE, and Cloud Run.
  • Write reusable Terraform modules and CI/CD pipelines for automated testing and deployment.
  • Administer IAM roles, Workload Identity Federation, and GCP monitoring and alerting tools.

Latitude is a Service-Disabled Veteran Owned federal IT company that helps government agencies modernize technology. It promotes a remote work-from-home culture and focuses on automating mundane tasks and investing in employee growth.

$54,000–$64,800/yr
Europe

  • Own identity, access, and IT operations across the SaaS and cloud stack, automating repetitive tasks.
  • Manage user lifecycle, handle support requests, and administer tools like Google Workspace, GitHub, and 1Password.
  • Automate processes using Terraform, Python, and Slack APIs, and support security operations and deployments.

Centrifuge is building the open infrastructure for tokenized real-world assets, having partnered with S&P Dow Jones Indices and crossed $1.7B in TVL. The company is a well-funded, remote-first team of over 20 people, backed by leading investors and live across more than ten blockchains.

Global

  • Contribute to platform and harness engineering, including CI/CD and developer tooling.
  • Build systems to reduce toil and maintain production infrastructure under conversational AI traffic.
  • Participate in on-call rotation and incident management to ensure platform uptime.

Replicant builds an AI-powered customer service platform that helps contact centers resolve requests and improve agent performance. The company is distributed, with a focus on ownership and collaboration, and serves Fortune 500 companies.