Source Job

US

  • Own core platform infrastructure including Terraform migration, CI/CD pipelines, and environment provisioning as part of a growing team.
  • Build and maintain tooling and abstractions that let product engineers ship and run code without solving infrastructure problems from scratch.
  • Advance observability foundation, standardize monitoring and alerting, and support GCP infrastructure scaling.

Terraform CI/CD GCP Go Kubernetes

20 jobs similar to Platform Engineer

Jobs ranked by similarity.

Romania

  • Architect, deploy, and manage highly available, fault-tolerant cloud infrastructure across Google Cloud Platform (GCP) and Google Kubernetes Engine (GKE).
  • Maintain and scale declarative infrastructure using Terraform across a multi-hundred-file estate, enforcing GitOps workflows with Atlantis.
  • Build, maintain, and optimize robust automated pipelines for continuous integration and delivery using GitHub Actions, Jenkins, and ArgoCD.

Point Wild helps customers monitor, manage, and protect against the risks associated with their identities and personal information in a digital world. Backed by WndrCo, Warburg Pincus and General Catalyst, Point Wild is a scrappy, nimble organization dedicated to creating the world’s most comprehensive portfolio of industry-leading cybersecurity solutions.

US

  • Design, build, and maintain robust, scalable, and secure infrastructure systems supporting Laurel's AI-driven platform.
  • Manage and optimize cloud infrastructure (AWS and Azure), Kubernetes orchestration, and CI/CD pipelines to increase deployment frequency and reliability.
  • Implement comprehensive observability, monitoring, and alerting to maintain system health and partner with engineering teams to optimize performance and cost-efficiency.

Laurel is an AI Time platform for professional services firms, automating work time capture and connecting time data to business outcomes for clients like EY and Crowell & Moring. The company comprises top AI, product, and engineering talent, is VC-backed by Google Ventures and IVP, and fosters an inclusive, ambitious culture.

$165,000–$200,000/yr
US

  • Design, build, and maintain cloud infrastructure on GCP and AWS using Terraform, optimizing CI/CD pipelines for rapid deployments.
  • Implement comprehensive observability including monitoring, logging, alerting, and distributed tracing to ensure platform health.
  • Establish and enforce security best practices, support AI/ML infrastructure, and build developer experience tooling.

Re:Build operates an advanced, end-to-end manufacturing platform that partners with industrial companies to bring products from concept to full-scale production. The company is guided by The Re:Build Way principles and aims to revitalize America's manufacturing base, creating meaningful jobs across the country.

Americas

  • Design and evolve cloud architecture on GCP, expressing it entirely as code with Terraform following GitOps principles.
  • Build and own CI/CD pipelines for IaC, including Policy-as-Code guardrails, drift detection, and progressive rollout.
  • Advance Platform-as-a-Product by building self-serve capabilities so engineers can provision what they need.

Alpaca is a US-headquartered global leader in agent-first brokerage infrastructure for stocks, ETFs, options, crypto, fixed income, and more. With over 400 globally distributed team members and $400 million in funding from top-tier investors, Alpaca is committed to open-source contributions and fostering a vibrant community.

Brazil 4w PTO

  • Own critical infrastructure across compute, networking, CI/CD, Kubernetes, and observability.
  • Manage Kubernetes environments and infrastructure-as-code with Terraform, improving developer experience and reducing operational friction.
  • Lead production incident response, influence architecture, and integrate AI-powered tools to boost engineering efficiency.

Jobgether is an AI-powered recruitment platform that connects candidates with global hiring companies. This role is with a partner company, a globally distributed technology organization offering a collaborative, informal culture and long-term opportunities.

Global

  • Set the technical direction and roadmap for Platform engineering, owning the reliability, performance, and cost of Traild's GCP infrastructure.
  • Improve developer experience by enhancing CI/CD pipelines, build tooling, observability, and reducing friction for other engineers.
  • Lead a small team of platform engineers, staying hands-on while growing and supporting your team members.

Traild is a high-growth SaaS company redefining how finance teams operate by combining AI, automation, and payments infrastructure for B2B finance. They are a rapidly growing global team with a strong culture, as evidenced by an eNPS score of 78.

UK Ireland Estonia Netherlands Sweden Israel Eastern Europe Portugal Unlimited PTO

  • Develop and evolve foundational software and services enabling product and development teams.\n- Architect, design, and implement Infrastructure as Code using Terraform.\n- Deploy, manage, and optimize Kubernetes clusters on GCP (GKE) and AWS (EKS).

DoiT is a global technology company that helps cloud-driven organizations leverage cloud for business growth and innovation through data, technology, and human expertise. They work with over 4,000 customers worldwide and foster a remote-first, entrepreneurial culture.

$120,000–$155,000/yr
Global

  • Own infrastructure as code across development, staging, and production environments
  • Build, maintain, and improve CI/CD pipelines for reliable and efficient deployments
  • Manage cloud infrastructure, establish scalable engineering practices, and lead incident response

CelebriOS is a software company building B2B SaaS products that help businesses make better decisions and streamline operations. The company has a remote-first working environment and a benefits package designed to support their team.

$150,000–$165,000/yr
US Unlimited PTO

  • Design and manage high-availability platforms using Kubernetes, Terraform, and Ansible with native-AI capabilities.
  • Develop and operate the observability stack: Grafana, Mimir, Loki, Tempo, and Prometheus on Kubernetes via GitLab CI/CD.
  • Build automation scripts in Python, maintain GitOps pipelines, and mentor mid-level engineers.

Flexential builds and operates critical IT platforms including observability, DevOps, and ITSM technologies. The company fosters a collaborative engineering culture and values diversity.

$150,000–$210,000/yr
Global Unlimited PTO

  • Own the design, development, and operation of infrastructure and build/release pipelines.
  • Deploy IaC and automation using Terraform, Ansible, Helm, and Go to support platform and customer requirements.
  • Collaborate closely with Product to drive roadmap direction and improve how users build and deliver software.

Manifest is on a mission to secure the global software and AI supply chain. Founded by alumni from the Department of Defense, CISA, and Palantir, it is a well-funded early-stage startup backed by leading investors and trusted by government and enterprise organizations.

US

  • Implement and manage the infrastructure stack to enable the engineering team to ship quickly and effectively.
  • Proactively identify and eliminate bottlenecks in the devops process to ensure optimal developer velocity.
  • Maintain Tempo chain reliability, validator infrastructure, and explorer reliability.

Tempo is a layer-1 blockchain purpose-built for stablecoins and real-world payments, born from Stripe and Paradigm. They are a team of crypto-optimists building infrastructure for onchain payments.

$150,000–$250,000/yr
US Europe Singapore

  • Own production uptime, latency, provisioning speed, infrastructure cost, and incident response for core platform services.
  • Build and maintain AWS infrastructure using Terraform, Kubernetes/EKS, Helm, Docker, EC2, CodeBuild, ECR, S3, IAM, networking, and secrets management.
  • Design and improve backend and platform systems for scale — capacity planning, autoscaling, queueing, backpressure, cleanup jobs, retries, and rollback paths.

A fast-growing AI/ML platform startup building infrastructure for training, evaluating, and aligning AI models within reinforcement learning environments. The engineering team of ~15 includes competitive programming medalists, serial AI startup founders, and researchers published at top venues.

Canada

  • Design and evolve highly available cloud architecture on Google Cloud using Terraform and GitOps.
  • Build and maintain secure CI/CD pipelines for Infrastructure-as-Code and develop self-service developer platforms.
  • Strengthen platform observability and apply SRE principles to improve reliability and operational maturity.

They are a financial technology company that provides production-critical infrastructure. They have a globally distributed team and a culture of autonomy and async-first collaboration.

$140,000–$175,000/yr
US

  • Lead and grow a team of platform engineers, coaching them on infrastructure and cloud challenges.
  • Drive the platform roadmap, balancing reliability, cost, security, and developer experience with AWS and Kubernetes.
  • Partner cross-functionally to align platform priorities with business goals and ensure system reliability.

PerfectServe is a leading provider of clinical communication and physician scheduling solutions in the health IT space. The company has 400+ employees and 30,000+ customers, with over $100 million in annual revenue, and has received multiple Best in KLAS awards.

Europe

  • Manage and troubleshoot complex distributed large-scale software systems
  • Build scalable, secure and reliable container-based infrastructure
  • Automate software delivery processes with CI/CD pipelines

Coinspaid Dev is the engineering brand behind the technology, infrastructure, and R&D expertise built within Coinspaid, focusing on advancing blockchain infrastructure engineering. With over 120 engineers and more than 11 years of industry experience, they bring together teams building distributed systems and blockchain infrastructure across 20+ blockchain networks.

Global

  • Own and scale cloud infrastructure including compute, networking, storage, and data systems.
  • Lead BYOC and private cloud deployments with infrastructure-as-code and GitOps foundations.
  • Establish reliability through service-level objectives, observability, and incident response processes.

A technology company builds a developer-focused platform with scalable cloud infrastructure. This is a remote-first opportunity with a small, autonomous engineering team operating in North America, LATAM, and Europe, offering high autonomy and ownership.

Latin America

  • Build and operate model and inference serving infrastructure, managing latency, throughput, autoscaling, and reliability for real-time and batch inference.
  • Own the ML deployment lifecycle: model registry, versioning, promotion workflows, rollout strategies, and safe rollback.
  • Operate agentic and LLM workloads in production, managing inference providers, gateways, quotas, guardrails, and graceful degradation under load.

ReadyOn is an AI-native Labor Operating System that redefines how enterprises manage frontline labor by matching workers to shifts in real time. Headquartered in San Francisco with over 100 employees, it grew revenue 8x year over year in 2025.

$140,000–$220,000/yr
North America LATAM Europe

  • Own and scale the cloud infrastructure behind our open-source platform: compute, networking, and the data layer.
  • Lead BYOC: turn customer-cloud deployments into a real product, with provisioning, upgrades, and observability that scale past bespoke work per deal.
  • Make reliability a product feature: meaningful SLOs, and an incident process people trust.

Nango is a developer infrastructure company that provides API access for agents and apps, enabling AI applications to connect to the real world through integrations. With over 400 paying customers and a team of 14 from top tech companies like AWS, GitHub, and Okta, they are a YC-backed, multi-million ARR company that values ownership and autonomy.

Global

  • Build and configure GKE clusters and Google Cloud project environments using Terraform.
  • Implement and maintain CI/CD pipelines and environment promotion workflows.
  • Configure and maintain Google Cloud-native monitoring, alerting, and logging.

We design, build, and scale AI-powered solutions that create real business impact. We are building a high-performance culture grounded in five values: Empowering Excellence, Collaborative Teamwork, Unsolicited Respect, Consistent Transparency, and Efficient Communication.

Argentina

  • Architect and maintain critical cloud platform components on AWS EKS with high availability and automated resilience.
  • Establish SRE standards including SLO/SLI tracking, error budget frameworks, and automated operational tooling.
  • Design and implement OpenTelemetry capture pipelines for telemetry data feeding downstream platforms.

Inflect is a US-based advisory and marketplace that revolutionizes how companies buy and sell digital infrastructure services. They operate with a focus on high-impact consulting and autonomous work.